Claude Fable 5.1 vs Llama 3.1 Instruct 405B
14 SHARED BENCHMARKSAcross 14 shared benchmarks, Claude Fable 5.1 scores higher on 13 and Llama 3.1 Instruct 405B on 1. The widest gap is AA Agentic Index, where Claude Fable 5.1 scores 58.2 against 6.3.
ANTHROPICVSMETA14 SHARED13–1 HEAD-TO-HEAD
Best tracked score per model per benchmark (default configuration; source-attributed). ↓ marks lower-is-better metrics. Open either model for its full surface, provenance and pricing.
Compare Claude Fable 5.1 withEVERY TRACKED PAIRING
▶Anthropic12 PAIRINGS
▶OpenAI11 PAIRINGS
▶Google9 PAIRINGS
▶Alibaba12 PAIRINGS
▶DeepSeek4 PAIRINGS
▶Moonshot4 PAIRINGS
▶Meta1 PAIRING
▶Z.ai2 PAIRINGS
▶MiniMax1 PAIRING
▶NVIDIA2 PAIRINGS
Compare Llama 3.1 Instruct 405B withEVERY TRACKED PAIRING
▶Anthropic12 PAIRINGS
▶OpenAI11 PAIRINGS
▶Google9 PAIRINGS
▶Alibaba12 PAIRINGS
▶DeepSeek4 PAIRINGS
▶Moonshot4 PAIRINGS
▶Meta1 PAIRING
▶Z.ai2 PAIRINGS
▶MiniMax1 PAIRING
▶NVIDIA2 PAIRINGS