VECTOR WIREAI INTELLIGENCE
NVDA$1,847+3.2%MSFT$512+1.1%GOOGL$199-0.4%META$728+2.7%AMD$184-1.2%TSM$212+0.6%PLTR$98+4.1%AI IDX4,821+1.9%
PKT
SEEDRefresh Models Deals Regulatory Sources

Gemini 2.5 Pro vs Gemini 3 Pro

GooglevsGoogle52 shared benchmarks844 head-to-head
BenchmarkGemini 2.5 ProGemini 3 Pro
AA Agentic Index7.252
AA Intelligence2740.6
AA-LCR6673
AA-Omniscience-14.315.3
AIME 202588.3100
arc_agi_14175
ARC-AGI-24.954
arena_elo14461485
arena_text_factuality14621481
arena_vision12461290
Artificial Analysis Coding Index46.746.5
browsecomp9.959.2
browsecomp_zh32.266.8
coding_arena_elo12241438
critpt2.69.1
Fortress54.941.7
frontiermath_tier_44.218.8
gdpval8.534.2
GPQA Diamond84.491.9
HLE22.545.8
HLE (with tools)28.445.8
hmmt_feb_202582.597.5
hmmt_nov_20258093.3
IFBench4970.4
LiveCodeBench82.792
matharena_visual_math_overall77.284.2
mathvision73.386.1
MathVista83.989.8
MMLU-Pro8690.1
MMMU-Pro74.981
mrcr9389.7
multichallenge53.665.7
OCRBench85.990.3
ocrbench_v259.363.4
OmniScience Accuracy3955.8
OmniScience Non-Hallucination12.610
PropensityBench7952.9
scicode4356.1
SEAL VISTA50.851.5
simplebench62.476.4
simpleqa50.872.1
simpleqa_verified5672.9
swe_bench_bash53.674.2
SWE-bench Verified67.278
Terminal-Bench Hard26.556.9
vectara_answer_rate99.199.4
vectara_avg_summary_length106.4101.9
vectara_factual_consistency9386.4
vectara_hallucination_rate713.6
Video-MME84.888.4
τ²-Bench5498
τ²-Bench Telecom (AA run)54.198

Best tracked score per model per benchmark (default configuration; source-attributed). ↓ marks lower-is-better metrics. Open either model for its full surface, provenance and pricing.