VECTOR WIREAI INTELLIGENCE
NVDA$1,847+3.2%MSFT$512+1.1%GOOGL$199-0.4%META$728+2.7%AMD$184-1.2%TSM$212+0.6%PLTR$98+4.1%AI IDX4,821+1.9%
PKT
SEEDRefresh Models Deals Regulatory Sources

Claude Opus 4.8 vs Gemini 3 Pro

AnthropicvsGoogle50 shared benchmarks3415 head-to-head
BenchmarkClaude Opus 4.8Gemini 3 Pro
AA Agentic Index49.452
AA Intelligence57.340.6
AA-LCR7373
AA-Omniscience28.815.3
aime_202610091.7
arc_agi_192.575
ARC-AGI-272.154
arena_elo14821485
arena_text_factuality14641481
arena_vision12801290
Artificial Analysis Coding Index74.346.5
browsecomp88.559.2
charxiv_rq80.581.4
coding_arena_elo15391438
critpt20.99.1
cybergym83.139.9
deepsearchqa_f193.163.2
Fortress18.241.7
frontiermath_tier_431.318.8
gdpval54.234.2
GDPval-AA (Elo)18901195
GPQA Diamond93.691.9
HLE57.945.8
HLE (with tools)57.945.8
hmmt_feb_202696.786.4
hmmt_nov_202596.593.3
IFBench62.270.4
imo_answer_bench83.583.3
livebench77.273.4
longbench_v269.168.2
matharena_visual_math_overall81.684.2
mathvision86.786.1
MCP Atlas83.670.3
MMMU-Pro78.981
MMVU79.278.9
OmniScience Accuracy48.855.8
OmniScience Non-Hallucination60.710
scicode53.556.1
screenspot_pro_no_tools87.972.7
simplebench64.876.4
simpleqa_verified39.572.9
SWE-bench Multilingual84.468.7
SWE-bench Pro69.243.3
SWE-bench Verified88.678
Terminal-Bench 2.074.654.2
Terminal-Bench Hard58.356.9
tool_decathlon59.936.4
toolathlon59.936.4
Video-MME8688.4
τ²-Bench Telecom (AA run)94.498

Best tracked score per model per benchmark (default configuration; source-attributed). ↓ marks lower-is-better metrics. Open either model for its full surface, provenance and pricing.