VECTOR WIREAI INTELLIGENCE
PKT
Refresh Models Deals Regulatory Sources

Questions this page answers

Which is better, Claude 3.5 Sonnet or Granite 4.2 3B?
Across 6 shared benchmarks, Claude 3.5 Sonnet scores higher on 5 and Granite 4.2 3B on 1. The widest gap is LiveCodeBench v6, where Granite 4.2 3B scores 69.7 against 37.2.

Claude 3.5 Sonnet vs Granite 4.2 3B

Across 6 shared benchmarks, Claude 3.5 Sonnet scores higher on 5 and Granite 4.2 3B on 1. The widest gap is LiveCodeBench v6, where Granite 4.2 3B scores 69.7 against 37.2.

AnthropicvsIBM6 shared benchmarks51 head-to-head
BenchmarkClaude 3.5 SonnetGranite 4.2 3B
GPQA Diamond67.254.8
LiveCodeBench v637.269.7
MMLU-Pro7867.8
RULER 128K93.855.3
RULER 64K95.267.5
scicode36.624.1

Best tracked score per model per benchmark (default configuration; source-attributed). ↓ marks lower-is-better metrics. Open either model for its full surface, provenance and pricing.