Skip to content
VECTOR WIREAI INTELLIGENCE
UTC
Pro Mode — raw scores, one benchmark under the lensofficial card harness
The field · 7 measured
60.384.2

7 models measured, most on the official card harness. Claude 3.5 Sonnet tops the board at 84.2.

7 measured·Lifecycle Updated Oct 8, 2026Score basis
#◆Vendor$/M in$/M out
Leader — ignites whiteField — fill cools with scoreBars show 60.3–84.2 · ◆ solid = first-party or better · outlined = arm's-length
01Claude 3.5 SonnetAnthropic84.2$3.0/M$15.0/M$0.107
02DeepSeek-V3DeepSeek79.7$0.24/M$0.90/M$0.007
03GPT-4oOpenAI72.9$5.0/M$15.0/M$0.137
04DeepSeek-V2.5DeepSeek71.6———
05Qwen2.5 Instruct 72BAlibaba65.4$0.47/M$0.49/M$0.007
06Llama 3.1 Instruct 405BMeta63.9———
07DeepSeek-V2DeepSeek60.3———
◆ solid = independently verified · aggregator attested · vendor attributed — outlined = cross-reference · source-attributed. Dimmed ◆ = confidence inferred. “+N alts” = alternate disclosures in the drawer with why-demoted. Prices per million tokens; $/point = input price ÷ score. Every row opens its model dossier.

Verification: 7 models scored · 0 independently verified · 4 vendor cross-reference · 3 vendor-reported · 0 with source disagreement. How these tiers are assigned