VECTOR WIREAI INTELLIGENCE
PKT
Refresh Models Deals Regulatory Sources

Questions this page answers

Which is better, Claude Sonnet 4-20250514 (no thinking) or Gemini 3 Pro?
Across 5 shared benchmarks, Claude Sonnet 4-20250514 (no thinking) scores higher on 3 and Gemini 3 Pro on 2. The widest gap is vectara_avg_summary_length, where Claude Sonnet 4-20250514 (no thinking) scores 145.8 against 101.9.

Claude Sonnet 4-20250514 (no thinking) vs Gemini 3 Pro

Across 5 shared benchmarks, Claude Sonnet 4-20250514 (no thinking) scores higher on 3 and Gemini 3 Pro on 2. The widest gap is vectara_avg_summary_length, where Claude Sonnet 4-20250514 (no thinking) scores 145.8 against 101.9.

AnthropicvsGoogle5 shared benchmarks32 head-to-head
BenchmarkClaude Sonnet 4-20250514 (no thinking)Gemini 3 Pro
arena_vision11761305
vectara_answer_rate98.699.4
vectara_avg_summary_length145.8101.9
vectara_factual_consistency89.786.4
vectara_hallucination_rate10.313.6

Best tracked score per model per benchmark (default configuration; source-attributed). ↓ marks lower-is-better metrics. Open either model for its full surface, provenance and pricing.