Mistral Medium 3.5 128B
Mistral Medium 3.5 128B is behind the leaders in long context, instruction following, coding, factuality, and multimodal tasks. Too few results yet to rate reasoning, agentic tasks, safety, math, or multilingual tasks.
Capability profile
Bars show the model's median result as a share of the leading model's, per capability. Select a row to see its results.
Too few results yet to rate Reasoning, Agentic, Safety, Math or Multilingual.
Price
$1.50input$7.50outputper million tokens
From Artificial Analysis · 3 providers tracked · All prices
Evidence
30results on28benchmarks
- 3 independently verified
- 19 aggregator
- 8 vendor-reported
From 5 sources · latest Oct 8, 2026 · How verification works
API features
Tool callingStructured outputsReasoning
As listed by OpenRouter
Mistral Medium 3.5 128B benchmark results
30 results on 28 benchmarks, grouped by capability. Each bar is the result as a share of the capability leader's; every result links to its source.
26.2% behind the leader1 of 3 ranked benchmarks measured
- 69.33Oct 8, 2026
29.8% behind the leader1 of 3 ranked benchmarks measured
- IFBench68.78Oct 8, 2026aa_ifbench
Show 1 more instruction following resultHide 1 instruction following result
- 69.00Oct 7, 2026
31.7% behind the leader5 of 10 ranked benchmarks measured
- 77.60Oct 7, 2026
- SciCode40.16Oct 8, 2026aa_scicode
- Terminal-Bench 2.150.56Oct 8, 2026terminalbenchV21
- Terminal-Bench Hard33.33Oct 8, 2026aa_terminalbench_hard
- 1265.20Jun 22, 2026
34.0% behind the leader2 of 4 ranked benchmarks measured
- AA-Omniscience · Accuracy24.67Oct 8, 2026omniscienceAccuracy
- AA-Omniscience · Non-hallucination18.41Oct 8, 2026omniscienceNonHallucination
34.2% behind the leader1 of 6 ranked benchmarks measured
- MMMU-Pro64.86Oct 8, 2026aa_mmmu_pro
Show 1 more multimodal resultHide 1 multimodal result
- 1222.21Sep 20, 2026
0 of 6 ranked benchmarks measured
- 0.00Oct 8, 2026
- GPQA Diamond74.85Oct 8, 2026gpqa
- Humanity's Last Exam13.76Oct 8, 2026aa_hle
0 of 7 ranked benchmarks measured
- 13.16Oct 8, 2026
- Terminal-Bench 4.00.00Oct 8, 2026
- τ-Bench V3 · Banking15.05Oct 8, 2026tauBanking
Show 1 more agentic resultHide 1 agentic result
- 48.60Oct 7, 2026
More results
Benchmarks outside the capability baskets. They are not ranked against a leader.
- AA Intelligence14.00Sep 20, 2026Artificial Analysis Intelligence Index
- AA Intelligence14.19Oct 8, 2026aa_intelligence_index
- τ²-Bench Telecom (AA run)94.15Oct 8, 2026aa_tau2
- -36.80Oct 8, 2026
- Artificial Analysis Coding Index46.90Sep 9, 2026aa_coding_index
- 9.35Sep 9, 2026
Show 5 more resultsHide 5 results
- 86.30Oct 7, 2026
- 66.90Oct 7, 2026
- 32.10Oct 7, 2026
- TauBench V3 - Telecom91.40Oct 7, 2026Tau3 Telecom
- τ³-Bench Banking13.40Oct 7, 2026Tau3 Banking
Mistral Medium 3.5 128B: common questions
Who makes Mistral Medium 3.5 128B?
Mistral Medium 3.5 128B is made by Mistral.
When was Mistral Medium 3.5 128B released?
Mistral Medium 3.5 128B was released on May 22, 2026, according to Mistral's own announcement.
What is Mistral Medium 3.5 128B good at?
Mistral Medium 3.5 128B is behind the leaders in long context, instruction following, coding, factuality, and multimodal tasks. Too few results yet to rate reasoning, agentic tasks, safety, math, or multilingual tasks.
How much does Mistral Medium 3.5 128B cost?
Mistral Medium 3.5 128B costs $1.50 per million input tokens and $7.50 per million output tokens, according to Artificial Analysis. We track its price at 3 providers. At a mix of three input tokens to one output token, it costs more than 75% of the 330 priced models we track.
How many benchmarks has Mistral Medium 3.5 128B been tested on?
We track 30 results for Mistral Medium 3.5 128B on 28 benchmarks from 5 sources, 3 of them independently verified. The latest was recorded on Oct 8, 2026.
Which API features does Mistral Medium 3.5 128B support?
OpenRouter lists tool calling, structured outputs, and reasoning for Mistral Medium 3.5 128B.
About this record
Where Mistral Medium 3.5 128B's numbers come from, and every name it appears under.
- Tracked since
- Apr 30, 2026
- Newest source mention
- Sep 3, 2026
Where the results come from
Verification: 30 scores · 3 independently verified · 19 aggregator-attributed · 8 vendor-reported. How these tiers are assigned
From 5 sources on 3 sites. Artificial Analysis supplies 20 of them; the 3 independently verified results come from 2 sites. Bars are coloured by trust tier.
- artificialanalysis.ai20
- api.llm-stats.com8
- datasets-server.huggingface.co2