Skip to content
VECTOR WIREAI INTELLIGENCE
UTC
Pro Mode — raw scores, one benchmark under the lensthird-party eval harness
The field · 12 measured
094.7

12 models measured, most on the third-party eval harness. MiniMax Text 01 tops the board at 94.7.

12 measured·Lifecycle Updated Oct 8, 2026Score basis
#◆Vendor$/M in$/M out
Leader — ignites whiteField — fill cools with scoreBars show 0–94.7 · ◆ solid = first-party or better · outlined = arm's-length
01MiniMax Text 01MiniMax94.7———
02Claude 3.5 SonnetAnthropic93.8$3.0/M$15.0/M$0.096
03Nemotron 3 Ultra 550B A55B BaseNVIDIA92.49———
04DeepSeek-V3.2-Exp-BaseDeepSeek91.88———
05Gemini 1.5 ProGoogle91.7———
06Kimi K2 BaseMoonshot88.61———
07Gemini 2.0 Flash ExpGoogle86———
08Granite 4.2 30B+1 altIBM81.38VENDOR$0.16/M$0.65/M$0.005
09Granite 4.2 8B+2 altsIBM71.41VENDOR$0.06/M$0.25/M$0.002
10Mistral Large 3 675B Base 2512Mistral55.77———
11Granite 4.2 3B+2 altsIBM55.3VENDOR$0.03/M$0.12/M$0.001
12GLM-4.5-BaseZ.ai0———
◆ solid = independently verified · aggregator attested · vendor attributed — outlined = cross-reference · source-attributed. Dimmed ◆ = confidence inferred. “+N alts” = alternate disclosures in the drawer with why-demoted. Prices per million tokens; $/point = input price ÷ score. Every row opens its model dossier.

Verification: 12 models scored · 0 independently verified · 7 vendor cross-reference · 5 vendor-reported · 0 with source disagreement. How these tiers are assigned