Sonnet 5.5
Sonnet 5.5 is capable in math and coding and behind the leaders in reasoning. Too few results yet to rate agentic tasks, safety, long context, multimodal tasks, multilingual tasks, instruction following, or factuality.
Capability profile
Bars show the model's median result as a share of the leading model's, per capability. Select a row to see its results.
Too few results yet to rate Agentic, Safety, Long Context, Multimodal, Multilingual, Instruction Following or Factuality.
Price
$2.00input$10.00outputper million tokens
From Anthropic's own price page · 3 providers tracked · All prices
Evidence
76results on58benchmarks
- 23 independently verified
- 25 vendor-reported
- 28 cross-referenced
From 16 sources · latest Oct 7, 2026 · How verification works
API features
Tool callingStructured outputsReasoning
As listed by OpenRouter
Research
2 papers reference Sonnet 5.5Sonnet 5.5 benchmark results
76 results on 58 benchmarks, grouped by capability. Each bar is the result as a share of the capability leader's; every result links to its source.
17.1% behind the leader3 of 5 ranked benchmarks measured
- LiveBench · Mathematics96.13Oct 7, 2026livebench_math@2026-06-25
- 88.77Sep 30, 2026
- 80.49Sep 30, 2026
Show 1 more math resultHide 1 math result
- LiveBench · Mathematics96.69Oct 7, 2026livebench_math@2026-06-25
21.5% behind the leader5 of 10 ranked benchmarks measured
- LiveBench · Coding91.37Oct 7, 2026livebench_coding@2026-06-25
- 1786.26Oct 2, 2026
- 90.30Oct 6, 2026
- 81.30Oct 6, 2026
- LiveBench · Agentic Coding56.31Oct 7, 2026livebench_agentic_coding@2026-06-25
Show 5 more coding resultsHide 5 coding results
- LiveBench · Agentic Coding39.34Oct 7, 2026livebench_agentic_coding@2026-06-25
- LiveBench · Coding88.87Oct 7, 2026livebench_coding@2026-06-25
- 1715.22Sep 30, 2026
- 90.30Oct 7, 2026
- 81.30Oct 7, 2026
26.6% behind the leader3 of 6 ranked benchmarks measured
- Humanity's Last Exam64.50Oct 7, 2026Humanity's Last Exam with tools
- LiveBench · Reasoning91.63Oct 7, 2026livebench_reasoning@2026-06-25
- 75.90Oct 2, 2026
Show 2 more reasoning resultsHide 2 reasoning results
- Humanity's Last Exam64.50Oct 7, 2026Humanity's Last Exam with tools
- LiveBench · Reasoning86.81Oct 7, 2026livebench_reasoning@2026-06-25
0 of 6 ranked benchmarks measured
- 1289.62Oct 3, 2026
0 of 3 ranked benchmarks measured
- LiveBench · Instruction Following56.79Oct 7, 2026livebench_instruction_following@2026-06-25
- LiveBench · Instruction Following70.52Oct 7, 2026livebench_instruction_following@2026-06-25
0 of 4 ranked benchmarks measured
- 46.50Sep 30, 2026
More results
Benchmarks outside the capability baskets. They are not ranked against a leader.
- livebench_data_analysis59.49Oct 7, 2026livebench_data_analysis@2026-06-25
- livebench_language77.98Oct 7, 2026livebench_language@2026-06-25
- livebench_language83.43Oct 7, 2026livebench_language@2026-06-25
- livebench_data_analysis78.62Oct 7, 2026livebench_data_analysis@2026-06-25
- 1471Oct 5, 2026
- AA Intelligence56.00Oct 3, 2026Artificial Analysis Intelligence Index
Show 47 more resultsHide 47 results
- ArXivMath without tools86.80Sep 28, 2026
- Child safety multi-turn appropriate response rate (API, without a system prompt)86.00Oct 7, 2026
- Child safety multi-turn appropriate response rate (Claude.ai)98.00Oct 7, 2026
- 71.00Oct 6, 2026
- Election integrity multi-turn (appropriate response rate) - API, without a system prompt86.00Oct 7, 2026
- Election integrity multi-turn (appropriate response rate) - Claude.ai86.00Oct 7, 2026
- Election integrity single-turn benign requests (refusal rate) - API, without a system prompt0.00Oct 7, 2026
- Election integrity single-turn benign requests (refusal rate) - Claude.ai0.00Oct 7, 2026
- Election integrity single-turn harmful requests (harmless rate) - API, without a system prompt100.00Oct 7, 2026
- Election integrity single-turn harmful requests (harmless rate) - Claude.ai100.00Oct 7, 2026
- FrontierCode 1.1 (Main) max46.20Oct 7, 2026
- FrontierCode 1.1 (Main) xhigh52.10Oct 7, 2026
- FrontierCode Extended59.10Oct 7, 2026
- FrontierCode Extended xhigh64.40Oct 7, 2026
- GDPval-AA 2.11840.00Oct 7, 2026GDPval-AA v2.1
- GDPval-AA 2.11844.00Sep 28, 2026GDPval-AA v2.1
- GDPval-AA 2.11840.00Oct 7, 2026GDPval-AA v2.1
- 92.10Oct 6, 2026
- 65.40Oct 6, 2026
- 69.20Oct 6, 2026
- HealthBench Professional length-adjusted69.20Oct 7, 2026HealthBench Professional (Length-adjusted)
- HLE (with tools)64.50Oct 6, 2026Humanity's Last Exam (with tools)
- 11.70Oct 6, 2026
- 91.60Oct 6, 2026
- 76.90Oct 6, 2026
- 65.60Oct 6, 2026
- OSWorld 2.183.90Oct 7, 2026
- OSWorld 2.180.00Sep 28, 2026
- OSWorld 2.1 (offline subset)83.90Oct 7, 2026
- OSWorld 2.1 partial80.10Oct 6, 2026OSWorld 2.1 (partial)
- OSWorld 2.1 partial score80.10Sep 28, 2026
- OSWorld 2.1 strict pass rate43.50Sep 28, 2026
- 79.70Oct 6, 2026
- 79.70Oct 7, 2026
- Suicide and self-harm multi-turn appropriate response rate (API, without a system prompt)60.00Oct 7, 2026
- Suicide and self-harm multi-turn appropriate response rate (Claude.ai)100.00Oct 7, 2026
- Suicide and self-harm single-turn benign requests refusal rate (API, without a system prompt)0.04Oct 7, 2026
- Suicide and self-harm single-turn benign requests refusal rate (Claude.ai)0.41Oct 7, 2026
- Suicide and self-harm single-turn requests posing potential risk harmless rate (API, without a system prompt)99.07Oct 7, 2026
- Suicide and self-harm single-turn requests posing potential risk harmless rate (Claude.ai)99.81Oct 7, 2026
- 54.30Oct 6, 2026
- 54.30Oct 7, 2026
- Terminal-Bench 4.070.60Oct 7, 2026
- Terminal-Bench 4.070.60Oct 7, 2026
- Terminal-Bench-Science 0.159.90Oct 6, 2026
- Terminal-Bench-Science 0.159.90Oct 7, 2026
- 77.80Oct 6, 2026
Sonnet 5.5: common questions
Who makes Sonnet 5.5?
Sonnet 5.5 is made by Anthropic.
When was Sonnet 5.5 released?
Sonnet 5.5 was released on Sep 28, 2026, according to Anthropic's own announcement.
What is Sonnet 5.5 good at?
Sonnet 5.5 is capable in math and coding and behind the leaders in reasoning. Too few results yet to rate agentic tasks, safety, long context, multimodal tasks, multilingual tasks, instruction following, or factuality.
How much does Sonnet 5.5 cost?
Sonnet 5.5 costs $2.00 per million input tokens and $10.00 per million output tokens, according to Anthropic's own price page. We track its price at 3 providers. At a mix of three input tokens to one output token, it costs more than 82% of the 328 priced models we track.
How many benchmarks has Sonnet 5.5 been tested on?
We track 76 results for Sonnet 5.5 on 58 benchmarks from 16 sources, 23 of them independently verified. The latest was recorded on Oct 7, 2026.
Which API features does Sonnet 5.5 support?
OpenRouter lists tool calling, structured outputs, and reasoning for Sonnet 5.5.
About this record
Where Sonnet 5.5's numbers come from, and every name it appears under.
- Tracked since
- Aug 10, 2026
- Newest source mention
- Oct 2, 2026
Where the results come from
Verification: 76 scores · 23 independently verified · 28 vendor cross-reference · 25 vendor-reported. How these tiers are assigned
From 16 sources on 9 sites. www-cdn.anthropic.com supplies 33 of them; the 23 independently verified results come from 6 sites. Bars are coloured by trust tier.
- www-cdn.anthropic.com33
- api.llm-stats.com16
- livebench.ai14
- anthropic.com4
- datasets-server.huggingface.co3
- epoch.ai3
- artificialanalysis.ai1
- lmarena.ai1
- simple-bench.com1
Also known as
How our sources name Sonnet 5.5 at each reasoning setting.
| Setting | Short form | API id |
|---|---|---|
| low | — | claude-sonnet-5-5-low |
| medium | — | claude-sonnet-5-5-medium |
| high | — | claude-sonnet-5-5-high claude-sonnet-5.5-high |
| xhigh | — | claude-sonnet-5-5-xhigh claude-sonnet-5.5-xhigh |
| max | claude sonnet 5.5 (max) | — |