VECTOR WIREAI INTELLIGENCE
NVDA$1,847+3.2%MSFT$512+1.1%GOOGL$199-0.4%META$728+2.7%AMD$184-1.2%TSM$212+0.6%PLTR$98+4.1%AI IDX4,821+1.9%
PKT
SEEDRefresh Models Deals Regulatory Sources

DeepSeek-V3.2 vs Gemini 3 Pro

DeepSeekvsGoogle65 shared benchmarks1054 head-to-head
BenchmarkDeepSeek-V3.2Gemini 3 Pro
AA Agentic Index39.852
AA Intelligence32.840.6
AA-LCR70.773
AA-Omniscience-22.515.3
AIME 202593.1100
AIME 2026 I92.790.6
aime_202695.191.7
AIME25 no tools89.396
apexAgents14.518.4
arc_agi_15775
ARC-AGI-2454
arena_text_factuality14331481
Arena-Hard (Creative Writing)88.893.6
Arena-Hard (Hard Prompt)53.472.6
Artificial Analysis Coding Index44.246.5
browsecomp51.459.2
BrowseComp (context management)67.667.6
browsecomp_with_context_manager67.659.2
browsecomp_zh6566.8
coding_arena_elo13681438
critpt2.99.1
cybergym17.339.9
deepsearchqa_f160.963.2
FinSearchCompT2&T359.149.9
frontiermath_tier_42.118.8
gdpval18.834.2
GPQA Diamond8491.9
HLE40.845.8
HLE (with tools)40.845.8
hmmt_feb_202592.597.5
hmmt_feb_202684.186.4
hmmt_nov_202590.293.3
IFBench6170.4
imo_answer_bench78.383.3
LiveCodeBench8692
LiveCodeBench v683.390.7
longbench_v259.868.2
MCP Atlas62.270.3
MMLU-Pro8690.1
mrcr55.589.7
Multi-SWE-Bench37.442.7
OctoCodingbench2622.9
OJ-Bench (cpp)54.768.5
OmniScience Accuracy3355.8
OmniScience Non-Hallucination17.310
scicode3956.1
Seal-049.545.5
simpleqa_verified27.572.9
swe_bench_bash6074.2
SWE-bench Multilingual70.268.7
SWE-bench Pro15.643.3
SWE-bench Verified73.178
SWE-Perf0.96.5
SWT-bench6279.7
Terminal-Bench 2.046.454.2
Terminal-Bench Hard35.656.9
tool_decathlon35.236.4
toolathlon35.236.4
vectara_answer_rate92.699.4
vectara_avg_summary_length62101.9
vectara_factual_consistency93.786.4
vectara_hallucination_rate6.313.6
WideSearch (item-f1)32.557
τ²-Bench9198
τ²-Bench Telecom (AA run)90.698

Best tracked score per model per benchmark (default configuration; source-attributed). ↓ marks lower-is-better metrics. Open either model for its full surface, provenance and pricing.