Gemini 3.1 Pro vs MiMo V2.5
32 SHARED BENCHMARKSAcross 32 shared benchmarks, Gemini 3.1 Pro scores higher on 26 and MiMo V2.5 on 6. The widest gap is CritPt, where Gemini 3.1 Pro scores 17.7 against 3.7. MiMo V2.5 is the cheaper of the two on tracked API pricing ($0.14 against $2.00 per million input tokens).
At a glance
Release dates per Artificial Analysis. Prices: Google's own price page for Gemini 3.1 Pro; Artificial Analysis for MiMo V2.5. ◆ = independently verified score.
Where each leadsby capability area · benchmark wins
Biggest gaps
Gemini 3.1 Pro pulls furthest ahead on
- CritPt17.7 vs 3.7
- SimpleQA Verified73.5 vs 16.1
- AA-Omniscience · Accuracy54.9 vs 16.8
MiMo V2.5 pulls furthest ahead on
- GDPVal24.3 vs 13.8
- AA-Omniscience · Non-hallucination68.1 vs 49.1
Every shared benchmark32 · grouped by area
Reasoning 3
Coding 6
Agentic 3
Factuality 3
Instruction Following 1
Long Context 1
Math 2
Multimodal 3
Other shared benchmarks 10
Each score is the one the model's own page shows — the most authoritative tracked result for that benchmark, on the benchmark's standard methodology (source-attributed). ↓ marks lower-is-better metrics; ◆ an independently verified score. Open either model for its full surface, provenance and pricing. Quoted rates are the price-setter row we currently track for each model — its direct or vendor-official listing where one exists (google-official, direct), otherwise the lowest tracked offer.