Skip to content
VECTOR WIREAI INTELLIGENCE
UTC
Pro Mode — raw scores, one benchmark under the lensthird-party eval harness
The field · 17 measured
9.00890.601

17 models measured, most on the third-party eval harness. Gemini 2.5 Pro tops the board at 90.601.

17 measured·Lifecycle Updated Oct 8, 2026Score basis
#◆Vendor$/M in$/M out
Leader — ignites whiteField — fill cools with scoreBars show 9.008–90.601 · ◆ solid = first-party or better · outlined = arm's-length
01Gemini 2.5 Pro+1 altGoogle90.601$1.3/M$10.0/M$0.062
02O3OpenAI89.817$2.0/M$8.0/M$0.056
03Qwen3 235B A22BAlibaba88.773$0.70/M$2.8/M$0.020
04EXAONE 4.0 32BLG AI88.251VENDOR———
05OpenReasoning-Nemotron-32BNVIDIA87.546VENDOR———
06O3 Mini+2 altsOpenAI87.467$1.1/M$4.4/M$0.031
07O4 Mini+2 altsOpenAI86.162VENDOR$1.1/M$4.4/M$0.032
08Grok 3 Mini ReasoningSpaceXAI85.901$0.30/M$0.50/M$0.005
09OpenCodeReasoning-Nemotron-1.1-32BNVIDIA85.692VENDOR———
10Gemini 2.5 Flash+1 altGoogle82.507$0.30/M$2.5/M$0.017
11Claude Opus 4Anthropic80.418VENDOR$15.0/M$75.0/M$0.560
12Claude Sonnet 4Anthropic75.979$3.0/M$15.0/M$0.118
13DeepSeek-V3DeepSeek53.499$0.24/M$0.90/M$0.011
14GPT-4oOpenAI32.063$5.0/M$15.0/M$0.312
15GPT-4 TurboOpenAI30.705$10.0/M$30.0/M$0.651
16GPT-4o miniOpenAI25.17$0.15/M$0.60/M$0.015
17Claude 3 HaikuAnthropic9.008$0.25/M$1.3/M$0.083
◆ solid = independently verified · aggregator attested · vendor attributed — outlined = cross-reference · source-attributed. Dimmed ◆ = confidence inferred. “+N alts” = alternate disclosures in the drawer with why-demoted. Prices per million tokens; $/point = input price ÷ score. Every row opens its model dossier.

Verification: 17 models scored · 17 independently verified · 0 with source disagreement. How these tiers are assigned