Mistral Medium 3.1 (Non-Reasoning)
Mistral Medium 3.1 (Non-Reasoning) is behind the leaders in multimodal tasks and coding. Too few results yet to rate reasoning, agentic tasks, safety, long context, math, multilingual tasks, instruction following, or factuality.
Capability profile
Bars show the model's median result as a share of the leading model's, per capability. Select a row to see its results.
Too few results yet to rate Reasoning, Agentic, Safety, Long Context, Math, Multilingual, Instruction Following or Factuality.
Price
$0.40input$2.00outputper million tokens
From mistralai · All prices
Evidence
26results on24benchmarks
- 7 independently verified
- 19 aggregator
From 5 sources · latest Oct 8, 2026 · How verification works
Mistral Medium 3.1 (Non-Reasoning) benchmark results
26 results on 24 benchmarks, grouped by capability. Each bar is the result as a share of the capability leader's; every result links to its source.
38.6% behind the leader1 of 6 ranked benchmarks measured
- MMMU-Pro54.16Oct 8, 2026aa_mmmu_pro
Show 2 more multimodal resultsHide 2 multimodal results
- 1172.29Sep 21, 2026
- 1159May 1, 2026
43.0% behind the leader3 of 10 ranked benchmarks measured
- SciCode32.41Oct 8, 2026aa_scicode
- Terminal-Bench Hard10.61Oct 8, 2026aa_terminalbench_hard
- Terminal-Bench 2.113.86Oct 8, 2026terminalbenchV21
0 of 6 ranked benchmarks measured
- 0.00Oct 8, 2026
- GPQA Diamond58.79Oct 8, 2026gpqa
- Humanity's Last Exam4.68Oct 8, 2026aa_hle
0 of 7 ranked benchmarks measured
- 0.00Oct 8, 2026
- Terminal-Bench 4.00.00Oct 8, 2026
- τ-Bench V3 · Banking8.25Oct 8, 2026tauBanking
0 of 3 ranked benchmarks measured
- 42.67Oct 8, 2026
0 of 3 ranked benchmarks measured
- IFBench39.80Oct 8, 2026aa_ifbench
0 of 4 ranked benchmarks measured
- 22.70Oct 8, 2026
- AA-Omniscience · Accuracy20.75Oct 8, 2026omniscienceAccuracy
- AA-Omniscience · Non-hallucination15.42Oct 8, 2026omniscienceNonHallucination
More results
Benchmarks outside the capability baskets. They are not ranked against a leader.
- vectara_avg_summary_length142.90Oct 8, 2026Average Summary Length (Words)
- vectara_answer_rate99.70Oct 8, 2026Answer Rate
- vectara_factual_consistency77.30Oct 8, 2026Factual Consistency Rate
- AA Intelligence15.00Sep 2, 2026Artificial Analysis Intelligence Index
- τ²-Bench Telecom (AA run)40.64Oct 8, 2026aa_tau2
- AA Intelligence9.18Oct 8, 2026aa_intelligence_index
Show 3 more resultsHide 3 results
- 3.10Sep 9, 2026
- -46.28Oct 8, 2026
- Artificial Analysis Coding Index20.50Sep 9, 2026aa_coding_index
Mistral Medium 3.1 (Non-Reasoning): common questions
Who makes Mistral Medium 3.1 (Non-Reasoning)?
Mistral Medium 3.1 (Non-Reasoning) is made by Mistral.
When was Mistral Medium 3.1 (Non-Reasoning) released?
Mistral Medium 3.1 (Non-Reasoning) was released on Aug 12, 2025, according to Artificial Analysis.
What is Mistral Medium 3.1 (Non-Reasoning) good at?
Mistral Medium 3.1 (Non-Reasoning) is behind the leaders in multimodal tasks and coding. Too few results yet to rate reasoning, agentic tasks, safety, long context, math, multilingual tasks, instruction following, or factuality.
How much does Mistral Medium 3.1 (Non-Reasoning) cost?
Mistral Medium 3.1 (Non-Reasoning) costs $0.40 per million input tokens and $2.00 per million output tokens, according to mistralai. At a mix of three input tokens to one output token, it costs more than 52% of the 331 priced models we track.
How many benchmarks has Mistral Medium 3.1 (Non-Reasoning) been tested on?
We track 26 results for Mistral Medium 3.1 (Non-Reasoning) on 24 benchmarks from 5 sources, 7 of them independently verified. The latest was recorded on Oct 8, 2026.
About this record
Where Mistral Medium 3.1 (Non-Reasoning)'s numbers come from, and every name it appears under.
- Tracked since
- Jun 19, 2026
- Newest source mention
- Jun 19, 2026
Where the results come from
Verification: 26 scores · 7 independently verified · 19 aggregator-attributed. How these tiers are assigned
From 5 sources on 4 sites. Artificial Analysis supplies 20 of them; the 7 independently verified results come from 4 sites. Bars are coloured by trust tier.
- artificialanalysis.ai20
- raw.githubusercontent.com4
- datasets-server.huggingface.co1
- lmarena.ai1