Skip to content
VECTOR WIREAI INTELLIGENCE
UTC

Claude Mythos 5 vs Claude Opus 5

Anthropic · released

Wins 2 of 4 areas

Agents · Images and charts

Anthropic · released

Wins 1 of 4 areas

Coding

Claude Mythos 5 wins more areas, narrowly.Claude Opus 5 is cheaper and better at coding.

Scores updated · 26 tests both models report · How we compare

Where each one wins

Tests won in each of the four areas where both have results. Each piece is one test, so a longer bar means more evidence; grey means the two scored within a point of each other.

Agents and images and charts rest on a single test each.

What it costs

Prices per million tokens, roughly 750,000 words. The bars show the cost of a million tokens read plus a million written.

Claude Mythos 5$10.00 to read · $50.00 to write$60Price from Anthropic's own price page
Claude Opus 5$5.00 to read · $25.00 to write$30Price from Anthropic's own price page

Claude Opus 5 costs 50% less for the same work.

The biggest differences

The three tests each model wins by the widest margin. Scores are out of 100.

Where Claude Mythos 5 pulls ahead

  • Reasoning about charts from research papersCharXiv (RQ)+5.2points aheadClaude Mythos 588.9Claude Opus 583.7
  • Very hard expert questions across many subjectsHumanity's Last Exam+2.9points aheadClaude Mythos 557.8Claude Opus 554.9
  • Completes tasks by operating a computer desktopOSWorld-Verified+1.6points aheadClaude Mythos 585Claude Opus 583.4

Where Claude Opus 5 pulls ahead

  • Fixes real GitHub issues in many programming languagesSWE-bench Multilingual+2.9points aheadClaude Opus 589.5Claude Mythos 586.6
  • Abstract visual puzzles that people can solveARC-AGI-2+1.2points aheadClaude Opus 590.4Claude Mythos 589.2
  • Command-line tasks in a real terminalTerminal-Bench 2.1+1.1points aheadClaude Opus 589.1Claude Mythos 588

Every test, side by side

All 26 tests both models report. The winning score is in its model's colour; marks a score checked independently.

CodingClaude Opus 5

Full coding ranking

AgentsClaude Mythos 5

Full agents ranking

ReasoningEven

Full reasoning ranking

Images and chartsClaude Mythos 5

Full images and charts ranking

Other results17 tests, not counted

Tests outside the eight areas. They are not counted above: several are summary scores built from other tests, or the same test under another name.

Questions people ask

Which is better, Claude Mythos 5 or Claude Opus 5?

Claude Mythos 5 wins two of the four areas where both have results: agents and images and charts. Claude Opus 5 wins coding, and costs 50% less. They are level on reasoning.

Which is better for coding?

Claude Opus 5. It wins 2 of the 4 coding tests both models report; Claude Mythos 5 wins none, and 2 are ties.

Which is cheaper?

Claude Mythos 5 costs $10.00 per million input tokens and $50.00 per million output tokens; Claude Opus 5 costs $5.00 and $25.00. That makes Claude Opus 5 about 50% cheaper for the same work.

How do you compare the two?

We use the 26 benchmark tests both models have published scores on. The verdict counts the 9 tests in the eight capability areas, and a gap under one point (ten on rating-style scales) counts as a tie. The other 17 are listed but not counted, because several are summary scores or repeat a test. Each score is the one shown on the model's own page.