VECTOR WIREAI INTELLIGENCE
UTC
Refresh Models Deals Regulatory Sources

Claude Opus 4.8 vs MiMo V2.6 Pro

19 SHARED BENCHMARKS

Across 19 shared benchmarks, Claude Opus 4.8 scores higher on 6 and MiMo V2.6 Pro on 13. The widest gap is AA-Omniscience, where Claude Opus 4.8 scores 28.8 against 8.4. MiMo V2.6 Pro is the cheaper of the two on tracked API pricing ($0.43 against $5.00 per million input tokens).

ANTHROPICVSXIAOMI19 SHARED613 HEAD-TO-HEADNEWEST SCORE ADDED

At a glance

MakerAnthropicXiaomi
Released
Price per 1M tokens input / output$5.00 / $25.00$0.43 / $0.87
Cost of 1M in + 1M out$30.00$1.30 23.1× less
Head-to-head of 19 shared benchmarks6 wins13 wins
Scores tracked independently verified211 3221 1

Release dates per Artificial Analysis. Prices: Artificial Analysis for Claude Opus 4.8; Artificial Analysis for MiMo V2.6 Pro. ◆ = independently verified score.

Where each leadsby capability area · benchmark wins

AreaClaude Opus 4.8WINSMiMo V2.6 Pro
Reasoning02MiMo V2.6 Pro leads 2 of 2 · widest: CritPt 20.9 vs 26.6
Coding02MiMo V2.6 Pro leads 2 of 2 · widest: SciCode 54.4 vs 60.9
Agentic12MiMo V2.6 Pro leads 2 of 3 · widest: Terminal-Bench 4.0 21.7 vs 34.8
Factuality20Claude Opus 4.8 leads 2 of 2 · widest: AA-Omniscience · Accuracy 48.8 vs 34.9
Long Context01MiMo V2.6 Pro leads 1 of 1 · widest: AA-LCR 77.7 vs 86.3

Biggest gaps

Claude Opus 4.8 pulls furthest ahead on

  1. AA-Omniscience · Accuracy48.8 vs 34.9

MiMo V2.6 Pro pulls furthest ahead on

  1. Terminal-Bench 4.034.8 vs 21.7
  2. CritPt26.6 vs 20.9
  3. GDPVal58.7 vs 46.9

Every shared benchmark19 · grouped by area

Reasoning 2

Coding 2

Agentic 3

GDPVal46.958.7
Terminal-Bench 4.021.734.8

Long Context 1

AA-LCR77.786.3

Other shared benchmarks 9

CyberGym78.894
JobBench48.462
Terminal-Bench 4.023.634.9

Each score is the one the model's own page shows — the most authoritative tracked result for that benchmark, on the benchmark's standard methodology (source-attributed). ↓ marks lower-is-better metrics; ◆ an independently verified score. Open either model for its full surface, provenance and pricing. Quoted rates are the price-setter row we currently track for each model — its direct or vendor-official listing where one exists (direct, direct), otherwise the lowest tracked offer.

Compare Claude Opus 4.8 withALL PAIRINGS →