VECTOR WIREAI INTELLIGENCE
PKT
Refresh Models Deals Regulatory Sources

Questions this page answers

Which is better, Claude 3.5 Sonnet or Gemma 3 1B?
Across 8 shared benchmarks, Claude 3.5 Sonnet scores higher on 8 and Gemma 3 1B on 0. The widest gap is MMLU-Pro, where Claude 3.5 Sonnet scores 78 against 14.7.

Claude 3.5 Sonnet vs Gemma 3 1B

Across 8 shared benchmarks, Claude 3.5 Sonnet scores higher on 8 and Gemma 3 1B on 0. The widest gap is MMLU-Pro, where Claude 3.5 Sonnet scores 78 against 14.7.

AnthropicvsGoogle8 shared benchmarks80 head-to-head
BenchmarkClaude 3.5 SonnetGemma 3 1B
bbh93.139.1
GPQA Diamond67.219.2
GSM8K96.962.8
humaneval93.741.5
ifeval90.180.2
LiveCodeBench32.81.9
MMLU-Pro7814.7
simpleqa28.42.2

Best tracked score per model per benchmark (default configuration; source-attributed). ↓ marks lower-is-better metrics. Open either model for its full surface, provenance and pricing.