Claude Mythos 5 vs Claude Mythos Preview
16 SHARED BENCHMARKSAcross 16 shared benchmarks, Claude Mythos 5 scores higher on 10 and Claude Mythos Preview on 6. The widest gap is RiemannBench, where Claude Mythos 5 scores 55 against 43.
ANTHROPICVSANTHROPIC16 SHARED10–6 HEAD-TO-HEADNEWEST SCORE ADDED
At a glance
MakerAnthropicAnthropic
Released—
Price per 1M tokens input / output$10.00 / $50.00—
Cost of 1M in + 1M out$60.00—
Head-to-head of 16 shared benchmarks10 wins6 wins
Scores tracked independently verified48 0 ◆25 0 ◆
Claude Mythos 5's release date per the vendor's own announcement. Prices: Anthropic's own price page for Claude Mythos 5. ◆ = independently verified score.
Where each leadsby capability area · benchmark wins
AreaClaude Mythos 5WINSClaude Mythos Preview
Biggest gaps
Claude Mythos 5 pulls furthest ahead on
- OSWorld-Verified85 vs 79.6
Claude Mythos Preview pulls furthest ahead on
- Humanity's Last Exam64.7 vs 57.8
- CharXiv (RQ)93.2 vs 88.9
Every shared benchmark16 · grouped by area
Reasoning 1
Coding 3
Agentic 1
Multimodal 1
Other shared benchmarks 10
Each score is the one the model's own page shows — the most authoritative tracked result for that benchmark, on the benchmark's standard methodology (source-attributed). ↓ marks lower-is-better metrics; ◆ an independently verified score. Open either model for its full surface, provenance and pricing.