Claude Opus 4.7 vs Claude Opus 5.5
32 SHARED BENCHMARKSAcross 32 shared benchmarks, Claude Opus 4.7 scores higher on 5 and Claude Opus 5.5 on 27. The widest gap is FrontierMath Tier 4, where Claude Opus 5.5 scores 95 against 31.7. Claude Opus 5.5 is the cheaper of the two on tracked API pricing ($4.00 against $5.00 per million input tokens).
At a glance
Release dates: Artificial Analysis for Claude Opus 4.7; the vendor's own announcement for Claude Opus 5.5. Prices: Artificial Analysis for Claude Opus 4.7; Anthropic's own price page for Claude Opus 5.5. ◆ = independently verified score.
Where each leadsby capability area · benchmark wins
Biggest gaps
Claude Opus 4.7 pulls furthest ahead on
- AA-Omniscience · Non-hallucination57.7 vs 41.4
- AA IT-Bench SRE46.7 vs 38.2
Claude Opus 5.5 pulls furthest ahead on
- FrontierMath Tier 495 vs 31.7
- CritPt31.7 vs 12
- GDPVal67.3 vs 41.9
Every shared benchmark32 · grouped by area
Reasoning 5
Coding 6
Agentic 2
Factuality 3
Instruction Following 1
Long Context 1
Math 3
Multimodal 1
Other shared benchmarks 10
Each score is the one the model's own page shows — the most authoritative tracked result for that benchmark, on the benchmark's standard methodology (source-attributed). ↓ marks lower-is-better metrics; ◆ an independently verified score. Open either model for its full surface, provenance and pricing. Quoted rates are the price-setter row we currently track for each model — its direct or vendor-official listing where one exists (direct, anthropic-official), otherwise the lowest tracked offer.