Skip to content
VECTOR WIREAI INTELLIGENCE
UTC
Pro Mode — raw scores, one benchmark under the lensofficial card harness
The field · 9 measured
10.878.8

9 models measured, most on the official card harness. DeepSeek-R1 tops the board at 78.8.

9 measured·Lifecycle Updated Oct 8, 2026Score basis
#◆Vendor$/M in$/M out
Leader — ignites whiteField — fill cools with scoreBars show 10.8–78.8 · ◆ solid = first-party or better · outlined = arm's-length
01DeepSeek-R1DeepSeek78.8$2.0/M$4.0/M$0.038
02Kimi K2 Instruct+3 altsMoonshot74.3$0.57/M$2.3/M$0.019
03O1 MiniOpenAI67.6———
04Claude Sonnet 4+1 altAnthropic60.43P$3.0/M$15.0/M$0.149
05Claude Opus 4+1 altAnthropic57.63P$15.0/M$75.0/M$0.781
06Qwen3 235B A22B+1 altAlibaba48.63P$0.70/M$2.8/M$0.036
07DeepSeek-V3+3 altsDeepSeek43.2$0.24/M$0.90/M$0.013
08Claude 3.5 SonnetAnthropic13.1$3.0/M$15.0/M$0.687
09GPT-4oOpenAI10.8$5.0/M$15.0/M$0.926
◆ solid = independently verified · aggregator attested · vendor attributed — outlined = cross-reference · source-attributed. Dimmed ◆ = confidence inferred. “+N alts” = alternate disclosures in the drawer with why-demoted. Prices per million tokens; $/point = input price ÷ score. Every row opens its model dossier.

Verification: 9 models scored · 0 independently verified · 3 vendor cross-reference · 6 vendor-reported · 0 with source disagreement. How these tiers are assigned