VECTOR WIREAI INTELLIGENCE
UTC
Refresh Models Deals Regulatory Sources

Claude 3.5 Sonnet vs IBM Granite 4.0 Small

6 SHARED BENCHMARKS

Across 6 shared benchmarks, Claude 3.5 Sonnet scores higher on 4 and IBM Granite 4.0 Small on 1, with 1 level. The widest gap is xstest, where Claude 3.5 Sonnet scores 95.6 against 80.6.

ANTHROPICVSIBM6 SHARED41 HEAD-TO-HEAD

Best tracked score per model per benchmark (default configuration; source-attributed). ↓ marks lower-is-better metrics. Open either model for its full surface, provenance and pricing.

Compare Claude 3.5 Sonnet withEVERY TRACKED PAIRING

Anthropic12 PAIRINGS
OpenAI11 PAIRINGS
Google9 PAIRINGS
Alibaba12 PAIRINGS
DeepSeek4 PAIRINGS
Moonshot4 PAIRINGS
Meta2 PAIRINGS
Z.ai2 PAIRINGS
MiniMax1 PAIRING
NVIDIA2 PAIRINGS