VECTOR WIREAI INTELLIGENCE
UTC
Refresh Models Deals Regulatory Sources

GPT-5.6 Sol vs MiMo V2.6 Pro

20 SHARED BENCHMARKS

Across 20 shared benchmarks, GPT-5.6 Sol scores higher on 10 and MiMo V2.6 Pro on 10. The widest gap is AA-Omniscience · Non-hallucination, where MiMo V2.6 Pro scores 59.4 against 8.1. MiMo V2.6 Pro is the cheaper of the two on tracked API pricing ($0.43 against $4.00 per million input tokens).

OPENAIVSXIAOMI20 SHARED1010 HEAD-TO-HEADNEWEST SCORE ADDED

At a glance

MakerOpenAIXiaomi
Released
Price per 1M tokens input / output$4.00 / $20.00$0.43 / $0.87
Cost of 1M in + 1M out$24.00$1.30 18.5× less
Head-to-head of 20 shared benchmarks10 wins10 wins
Scores tracked independently verified222 3121 1

Release dates per Artificial Analysis. Prices: Artificial Analysis for GPT-5.6 Sol; Artificial Analysis for MiMo V2.6 Pro. ◆ = independently verified score.

Where each leadsby capability area · benchmark wins

AreaGPT-5.6 SolWINSMiMo V2.6 Pro
Reasoning20GPT-5.6 Sol leads 2 of 2 · widest: CritPt 32.3 vs 26.6
Coding02MiMo V2.6 Pro leads 2 of 2 · widest: SciCode 57.1 vs 60.9
Agentic21GPT-5.6 Sol leads 2 of 3 · widest: Terminal-Bench 4.0 39.9 vs 34.8
Factuality11Even, 1–1 of 2
Long Context01MiMo V2.6 Pro leads 1 of 1

Biggest gaps

GPT-5.6 Sol pulls furthest ahead on

  1. AA-Omniscience · Accuracy59.4 vs 34.9
  2. CritPt32.3 vs 26.6
  3. Terminal-Bench 4.039.9 vs 34.8

MiMo V2.6 Pro pulls furthest ahead on

  1. AA-Omniscience · Non-hallucination59.4 vs 8.1
  2. GDPVal58.7 vs 54.4
  3. SciCode60.9 vs 57.1

Every shared benchmark20 · grouped by area

Reasoning 2

Coding 2

Agentic 3

GDPVal54.458.7
Terminal-Bench 4.039.934.8

Long Context 1

AA-LCR8486.3

Other shared benchmarks 10

CyberGym84.594
GDPval-AA 2.115881673
JobBench45.462
Terminal-Bench 4.052.334.9

Each score is the one the model's own page shows — the most authoritative tracked result for that benchmark, on the benchmark's standard methodology (source-attributed). ↓ marks lower-is-better metrics; ◆ an independently verified score. Open either model for its full surface, provenance and pricing. Quoted rates are the price-setter row we currently track for each model — its direct or vendor-official listing where one exists (direct, direct), otherwise the lowest tracked offer.

Compare GPT-5.6 Sol withALL PAIRINGS →