Skip to content
VECTOR WIREAI INTELLIGENCE
UTC

Claude Opus 4.1

Score basis

Claude Opus 4.1 is capable in long context; and behind the leaders in coding, multimodal tasks, instruction following, and math. Too few results yet to rate reasoning, agentic tasks, safety, multilingual tasks, or factuality.

Capability profile

Bars show the model's median result as a share of the leading model's, per capability. Select a row to see its results.

Capable1 capability
  1. Long Context−17.8%1 of 3
Limited4 capabilities
  1. Coding−31.5%4 of 10
  2. Multimodal−33.4%1 of 6
  3. Instruction Following−34.8%2 of 3
  4. Math−56.0%2 of 5
Not rated5 capabilities

Too few results yet to rate Reasoning, Agentic, Safety, Multilingual or Factuality.

Price

$15.00input$75.00outputper million tokens

From Anthropic's own price page · 3 providers tracked · All prices

Costs more than 97% of 330 priced models · 3:1 input-to-output blend, log scale

Evidence

45results on40benchmarks

  • 17 independently verified
  • 12 aggregator
  • 16 vendor-reported

From 15 sources · latest Oct 8, 2026 · How verification works

API features

Tool callingStructured outputsReasoning

As listed by OpenRouter

Claude Opus 4.1 benchmark results

45 results on 40 benchmarks, grouped by capability. Each bar is the result as a share of the capability leader's; every result links to its source.

Long Context

CapableFull long context ranking

17.8% behind the leader1 of 3 ranked benchmarks measured

Coding

LimitedFull coding ranking

31.5% behind the leader4 of 10 ranked benchmarks measured

Multimodal

LimitedFull multimodal ranking

33.4% behind the leader1 of 6 ranked benchmarks measured

Instruction Following

LimitedFull instruction following ranking

34.8% behind the leader2 of 3 ranked benchmarks measured

56.0% behind the leader2 of 5 ranked benchmarks measured

Reasoning

Not enough dataFull reasoning ranking

0 of 6 ranked benchmarks measured

Show 3 more reasoning resultsHide 3 reasoning results

Agentic

Not enough dataFull agentic ranking

0 of 7 ranked benchmarks measured

Factuality

Not enough dataFull factuality ranking

0 of 4 ranked benchmarks measured

More results

Benchmarks outside the capability baskets. They are not ranked against a leader.

Show 19 more resultsHide 19 results

Claude Opus 4.1: common questions

Who makes Claude Opus 4.1?

Claude Opus 4.1 is made by Anthropic.

When was Claude Opus 4.1 released?

Claude Opus 4.1 was released on Aug 5, 2025, according to Artificial Analysis.

What is Claude Opus 4.1 good at?

Claude Opus 4.1 is capable in long context; and behind the leaders in coding, multimodal tasks, instruction following, and math. Too few results yet to rate reasoning, agentic tasks, safety, multilingual tasks, or factuality.

How much does Claude Opus 4.1 cost?

Claude Opus 4.1 costs $15.00 per million input tokens and $75.00 per million output tokens, according to Anthropic's own price page. We track its price at 3 providers. At a mix of three input tokens to one output token, it costs more than 97% of the 330 priced models we track.

How many benchmarks has Claude Opus 4.1 been tested on?

We track 45 results for Claude Opus 4.1 on 40 benchmarks from 15 sources, 17 of them independently verified. The latest was recorded on Oct 8, 2026.

Which API features does Claude Opus 4.1 support?

OpenRouter lists tool calling, structured outputs, and reasoning for Claude Opus 4.1.

About this record

Where Claude Opus 4.1's numbers come from, and every name it appears under.

Tracked since
Apr 27, 2026
Newest source mention
Aug 31, 2026

Where the results come from

Verification: 45 scores · 17 independently verified · 12 aggregator-attributed · 16 vendor-reported. How these tiers are assigned

From 15 sources on 9 sites. Artificial Analysis supplies 13 of them; the 17 independently verified results come from 7 sites. Bars are coloured by trust tier.

  • artificialanalysis.ai13
  • www-cdn.anthropic.com9
  • api.llm-stats.com7
  • raw.githubusercontent.com7
  • epoch.ai4
  • lmarena.ai2
  • datasets-server.huggingface.co1
  • labs.scale.com1
  • simple-bench.com1

Also known as

claude 4.1 opusclaude 4.1 opus (inspect)claude 4.1 opus (non-reasoning)claude 4.1 opus (reasoning)Claude Opus 4-1-20250805 ThinkingClaude Opus 4-1-20250805 Thinking 16KClaude Opus 4.1:batchclaude-4-1-opusclaude-4-1-opus-thinkingclaude-opus-4-1-20250805claude's opus 4.1opus 4.1terminus 2 + claude opus 4.1