Skip to content
VECTOR WIREAI INTELLIGENCE
UTC
Pro Mode — raw scores, one benchmark under the lensthird-party eval harness
The field · 5 measured
48.3383.33

5 models measured, most on the third-party eval harness. Gemini 3 Deep Think tops the board at 83.33.

5 measured·Lifecycle Updated Oct 8, 2026Score basis
#◆Vendor$/M in$/M out
Leader — ignites whiteField — fill cools with scoreBars show 48.33–83.33 · ◆ solid = first-party or better · outlined = arm's-length
01Gemini 3 Deep ThinkGoogle83.33———
02GPT-5 ProOpenAI77.5$15.0/M$120.0/M$0.871
03Gemini 3 ProGoogle75.83$2.0/M$12.0/M$0.092
04DeepSeek-V3.2-SpecialeDeepSeek59.17$0.29/M$0.43/M$0.006
05GPT-5.1OpenAI48.33$1.3/M$10.0/M$0.116
◆ solid = independently verified · aggregator attested · vendor attributed — outlined = cross-reference · source-attributed. Dimmed ◆ = confidence inferred. “+N alts” = alternate disclosures in the drawer with why-demoted. Prices per million tokens; $/point = input price ÷ score. Every row opens its model dossier.

Verification: 5 models scored · 5 independently verified · 0 with source disagreement. How these tiers are assigned