VECTOR WIREAI INTELLIGENCE
PKT
Refresh Models Deals Regulatory Sources

Questions this page answers

Which is better, Claude Sonnet 4-20250514 (no thinking) or Gemini 2.5 Flash Lite?
Across 6 shared benchmarks, Claude Sonnet 4-20250514 (no thinking) scores higher on 2 and Gemini 2.5 Flash Lite on 4. The widest gap is vectara_hallucination_rate, where Gemini 2.5 Flash Lite scores 3.3 against 10.3.

Claude Sonnet 4-20250514 (no thinking) vs Gemini 2.5 Flash Lite

Across 6 shared benchmarks, Claude Sonnet 4-20250514 (no thinking) scores higher on 2 and Gemini 2.5 Flash Lite on 4. The widest gap is vectara_hallucination_rate, where Gemini 2.5 Flash Lite scores 3.3 against 10.3.

AnthropicvsGoogle6 shared benchmarks24 head-to-head
BenchmarkClaude Sonnet 4-20250514 (no thinking)Gemini 2.5 Flash Lite
aider_polyglot56.426.7
arena_vision11761188
vectara_answer_rate98.699.5
vectara_avg_summary_length145.895.7
vectara_factual_consistency89.796.7
vectara_hallucination_rate10.33.3

Best tracked score per model per benchmark (default configuration; source-attributed). ↓ marks lower-is-better metrics. Open either model for its full surface, provenance and pricing.