VECTOR WIREAI INTELLIGENCE
UTC
Refresh Models Deals Regulatory Sources

GPT-5.5 vs MiMo V2.5 Pro

37 SHARED BENCHMARKS

Across 37 shared benchmarks, GPT-5.5 scores higher on 37 and MiMo V2.5 Pro on 0. The widest gap is Program Bench, where GPT-5.5 scores 70.8 against 12.5. MiMo V2.5 Pro is the cheaper of the two on tracked API pricing ($0.43 against $5.00 per million input tokens).

OPENAIVSXIAOMI37 SHARED370 HEAD-TO-HEADNEWEST SCORE ADDED

At a glance

MakerOpenAIXiaomi
Released
Price per 1M tokens input / output$5.00 / $30.00$0.43 / $0.87
Cost of 1M in + 1M out$35.00$1.30 26.9× less
Head-to-head of 37 shared benchmarks37 wins0 wins
Scores tracked independently verified204 35101 4

Release dates per Artificial Analysis. Prices: Artificial Analysis for GPT-5.5; Artificial Analysis for MiMo V2.5 Pro. ◆ = independently verified score.

Where each leadsby capability area · benchmark wins

AreaGPT-5.5WINSMiMo V2.5 Pro
Reasoning30GPT-5.5 leads 3 of 3 · widest: Humanity's Last Exam 45 vs 14.8
Coding50GPT-5.5 leads 5 of 5 · widest: Terminal-Bench Hard 59.8 vs 35.6
Agentic50GPT-5.5 leads 5 of 5 · widest: τ-Bench V3 · Banking 36.7 vs 9.9
Factuality30GPT-5.5 leads 3 of 3 · widest: SimpleQA Verified 63 vs 16.1
Instruction Following10GPT-5.5 leads 1 of 1 · widest: IFBench 71.6 vs 42.7
Long Context10GPT-5.5 leads 1 of 1 · widest: AA-LCR 84.3 vs 41.7
Math20GPT-5.5 leads 2 of 2 · widest: HMMT Feb. 2026 96.7 vs 82.6
Multimodal20GPT-5.5 leads 2 of 2 · widest: MMMU-Pro 81.1 vs 77.9

Biggest gaps

GPT-5.5 pulls furthest ahead on

  1. SimpleQA Verified63 vs 16.1
  2. τ-Bench V3 · Banking36.7 vs 9.9
  3. Humanity's Last Exam45 vs 14.8

MiMo V2.5 Pro pulls furthest ahead on

No ratified-area lead of 3 points or more.

Every shared benchmark37 · grouped by area

Reasoning 3

Coding 5

BenchmarkGPT-5.5MARGINMiMo V2.5 Pro
LMArena · WebDev1458 1437
SciCode56.150.6

Agentic 5

BenchmarkGPT-5.5MARGINMiMo V2.5 Pro
GDPVal40.530.4
Terminal-Bench 4.09.10

Instruction Following 1

BenchmarkGPT-5.5MARGINMiMo V2.5 Pro
IFBench71.642.7

Long Context 1

BenchmarkGPT-5.5MARGINMiMo V2.5 Pro
AA-LCR84.341.7

Math 2

BenchmarkGPT-5.5MARGINMiMo V2.5 Pro
AIME 202698.393.6

Multimodal 2

BenchmarkGPT-5.5MARGINMiMo V2.5 Pro
LMArena · Vision1294 1247
MMMU-Pro81.177.9

Each score is the one the model's own page shows — the most authoritative tracked result for that benchmark, on the benchmark's standard methodology (source-attributed). ↓ marks lower-is-better metrics; ◆ an independently verified score. Open either model for its full surface, provenance and pricing. Quoted rates are the price-setter row we currently track for each model — its direct or vendor-official listing where one exists (direct, direct), otherwise the lowest tracked offer.

Compare GPT-5.5 withALL PAIRINGS →

Compare MiMo V2.5 Pro withALL PAIRINGS →