GLM 5V Turbo vs MiMo V2 Omni
Wins 2 of 7 areas
Images and charts · Following instructions
Wins 2 of 7 areas
Reasoning · Long documents
The two are evenly matched.GLM 5V Turbo is better at images and charts and following instructions; MiMo V2 Omni at reasoning and long documents.
Scores updated · 19 tests both models report · How we compare
Where each one wins
Tests won in each of the seven areas where both have results. Each piece is one test, so a longer bar means more evidence; grey means the two scored within a point of each other.
- Images and chartsUnderstanding pictures, charts and video20GLM 5V Turbo2 of 2 tests
- Following instructionsDoing exactly what it is asked10GLM 5V Turbo1 of 1 test
- ReasoningHard problems that need careful thinking02MiMo V2 Omni2 of 3 tests · 1 tie
- Long documentsFinding answers in very long texts01MiMo V2 Omni1 of 1 test
- CodingWriting and fixing software11Even1 each
- AgentsCarrying out multi-step tasks on its own00Even0 each · 1 tie
- FactsGetting facts right instead of making them up11Even1 each
Agents, long documents and following instructions rest on a single test each.
What it costs
Prices per million tokens, roughly 750,000 words. The bars show the cost of a million tokens read plus a million written.
MiMo V2 Omni costs 54% less for the same work.
The biggest differences
The three tests each model wins by the widest margin. Scores are out of 100.
Where GLM 5V Turbo pulls ahead
- Answers hard knowledge questions correctlyAA-Omniscience · Accuracy+10points ahead
- Follows unfamiliar, precisely checkable instructionsIFBench+7.6points ahead
- Code for real scientific research problemsSciCode+6.8points ahead
Where MiMo V2 Omni pulls ahead
- Avoids making up answers it doesn't knowAA-Omniscience · Non-hallucination+19.9points ahead
- Very hard expert questions across many subjectsHumanity's Last Exam+5points ahead
- Reasons across sets of long documentsAA-LCR+4.7points ahead
Every test, side by side
All 19 tests both models report. The winning score is in its model's colour; marks a score checked independently.
CodingEven
- SciCodeGLM 5V Turbo by 6.843.536.7+6.8
- Terminal-Bench HardMiMo V2 Omni by 2.232.634.8+2.2
ReasoningMiMo V2 Omni
- Humanity's Last ExamMiMo V2 Omni by 517.122.1+5
- GPQA DiamondMiMo V2 Omni by 1.980.982.8+1.9
- CritPttie0.61.1tie
FactsEven
- AA-Omniscience · Non-hallucinationMiMo V2 Omni by 19.931.251.1+19.9
- AA-Omniscience · AccuracyGLM 5V Turbo by 1029.319.3+10
Images and chartsGLM 5V Turbo
- LMArena · VisionGLM 5V Turbo by 36 rating points12641228+36 rating
- MMMU-ProGLM 5V Turbo by 2.972.869.9+2.9
Following instructionsGLM 5V Turbo
- IFBenchGLM 5V Turbo by 7.661.153.5+7.6
Other results7 tests, not counted
Tests outside the eight areas. They are not counted above: several are summary scores built from other tests, or the same test under another name.
- Claw Eval (pass@3)GLM 5V Turbo by 20.27554.8+20.2
- τ²-Bench Telecom (AA run)GLM 5V Turbo by 7.398.591.2+7.3
- AA Agentic IndexGLM 5V Turbo by 2.561.158.6+2.5
- AA-Omnisciencetie-19.3-20.1tie
- Artificial Analysis Coding Indextie36.235.5tie
- PinchBenchtie80.781.2tie
- AA Intelligencetie23.523.9tie
Questions people ask
Which is better, GLM 5V Turbo or MiMo V2 Omni?
GLM 5V Turbo and MiMo V2 Omni each win two of the seven areas where both have results. GLM 5V Turbo wins images and charts and following instructions; MiMo V2 Omni wins reasoning and long documents. MiMo V2 Omni costs 54% less. They are level on coding, agents and facts.
Which is better for coding?
Neither. They win 1 coding test each of the 2 both models report.
Which is cheaper?
GLM 5V Turbo costs $1.20 per million input tokens and $4.00 per million output tokens; MiMo V2 Omni costs $0.40 and $2.00. That makes MiMo V2 Omni about 54% cheaper for the same work.
How do you compare the two?
We use the 19 benchmark tests both models have published scores on. The verdict counts the 12 tests in the eight capability areas, and a gap under one point (ten on rating-style scales) counts as a tie. The other 7 are listed but not counted, because several are summary scores or repeat a test. Each score is the one shown on the model's own page.