Claude Haiku 4.5
Claude Haiku 4.5 is capable in long context; and behind the leaders in multimodal tasks, coding, and instruction following. Too few results yet to rate reasoning, agentic tasks, safety, math, multilingual tasks, or factuality.
Capability profile
Bars show the model's median result as a share of the leading model's, per capability. Select a row to see its results.
Too few results yet to rate Reasoning, Agentic, Safety, Math, Multilingual or Factuality.
Price
$1.00input$5.00outputper million tokens
From Anthropic's own price page · 4 providers tracked · All prices
Evidence
81results on57benchmarks
- 27 independently verified
- 34 aggregator
- 19 vendor-reported
- 1 cross-referenced
From 22 sources · latest Oct 8, 2026 · How verification works
API features
Tool callingStructured outputsReasoning
As listed by OpenRouter
Claude Haiku 4.5 benchmark results
81 results on 57 benchmarks, grouped by capability. Each bar is the result as a share of the capability leader's; every result links to its source.
19.8% behind the leader1 of 3 ranked benchmarks measured
- 74.33Oct 8, 2026
Show 1 more long context resultHide 1 long context result
- 49.67Oct 8, 2026
37.1% behind the leader1 of 6 ranked benchmarks measured
- MMMU-Pro58.55Oct 8, 2026aa_mmmu_pro
Show 1 more multimodal resultHide 1 multimodal result
- MMMU-Pro55.14Oct 8, 2026aa_mmmu_pro
38.6% behind the leader7 of 10 ranked benchmarks measured
- 73.30Oct 7, 2026
- 67.40Oct 7, 2026
- SciCode42.25Oct 8, 2026aa_scicode
- Terminal-Bench 2.144.19Oct 8, 2026terminalbenchV21
- 1328.64May 22, 2026
- 39.45Oct 8, 2026
- Terminal-Bench Hard27.27Oct 8, 2026aa_terminalbench_hard
Show 2 more coding resultsHide 2 coding results
- SciCode34.38Sep 4, 2026aa_scicode
- 64.70May 1, 2026
39.2% behind the leader2 of 3 ranked benchmarks measured
- IFBench54.29Oct 8, 2026aa_ifbench
- 50.49Oct 8, 2026
Show 1 more instruction following resultHide 1 instruction following result
- IFBench42.04Oct 8, 2026aa_ifbench
0 of 6 ranked benchmarks measured
- 4.03Sep 22, 2026
- 2.78Sep 22, 2026
- 1.67Sep 22, 2026
Show 8 more reasoning resultsHide 8 reasoning results
- 1.25Sep 22, 2026
- 0.00Oct 8, 2026
- GPQA Diamond64.65Oct 8, 2026gpqa
- GPQA Diamond67.17Oct 8, 2026gpqa
- GPQA Diamond73.00Oct 7, 2026GPQA
- Humanity's Last Exam10.38Oct 8, 2026aa_hle
- Humanity's Last Exam4.22Oct 8, 2026aa_hle
- Humanity's Last Exam18.70Oct 7, 2026Humanity's Last Exam with tools
0 of 7 ranked benchmarks measured
- 40.20Oct 8, 2026
- 27.31Oct 8, 2026
- 11.76Oct 8, 2026
Show 4 more agentic resultsHide 4 agentic results
- 31.80Jun 15, 2026
- OSWorld-Verified50.70Oct 7, 2026OSWorld
- Terminal-Bench 4.00.00Oct 8, 2026
- τ-Bench V3 · Banking9.28Oct 8, 2026tauBanking
0 of 4 ranked benchmarks measured
- 13.20May 20, 2026
- 9.80May 2, 2026
- AA-Omniscience · Accuracy14.43Oct 8, 2026omniscienceAccuracy
Show 3 more factuality resultsHide 3 factuality results
- AA-Omniscience · Accuracy18.02Oct 8, 2026omniscienceAccuracy
- AA-Omniscience · Non-hallucination72.70Oct 8, 2026omniscienceNonHallucination
- AA-Omniscience · Non-hallucination74.29Oct 8, 2026omniscienceNonHallucination
More results
Benchmarks outside the capability baskets. They are not ranked against a leader.
- 47.67Sep 22, 2026
- 37.33Sep 22, 2026
- 25.50Sep 22, 2026
- 16.83Sep 22, 2026
- 98.75Sep 7, 2026
- 93.22Sep 7, 2026
Show 35 more resultsHide 35 results
- 10.31Sep 9, 2026
- 32.59Jun 18, 2026
- AA Intelligence15.41Oct 8, 2026aa_intelligence_index
- AA Intelligence16.88Oct 8, 2026aa_intelligence_index
- -7.57Oct 8, 2026
- -4.37Oct 8, 2026
- 80.70Oct 7, 2026
- 93.15Sep 7, 2026
- 96.88Sep 7, 2026
- 14.33May 10, 2026
- Artificial Analysis Coding Index43.89Sep 9, 2026aa_coding_index
- 29.64Jun 18, 2026
- 92.80Sep 7, 2026
- 54.70Jul 29, 2026
- 52.53Jul 29, 2026
- 31.01Oct 7, 2026
- 23.00Aug 25, 2026
- frontiermath_tier_4_v12.08May 20, 2026frontiermath_tier_4
- GDPval-AA 2.1735.00Oct 7, 2026GDPval-AA v2.1
- 95.88Sep 7, 2026
- HealthBench Professional length-adjusted32.20Oct 7, 2026HealthBench Professional (Length-adjusted)
- 83.00Oct 7, 2026
- OSWorld 2.115.70Oct 7, 2026
- OSWorld 2.1 (offline subset)15.70Oct 7, 2026
- 66.60May 1, 2026
- 19.80Oct 7, 2026
- 63.60Oct 7, 2026
- 41.00Sep 23, 2026
- Terminal-Bench 4.00.00Oct 7, 2026
- vectara_answer_rate99.50May 2, 2026Answer Rate
- vectara_avg_summary_length115.10May 2, 2026Average Summary Length (Words)
- vectara_factual_consistency90.20May 2, 2026Factual Consistency Rate
- τ²-Bench (Retail)83.20Oct 7, 2026Tau2 Retail
- τ²-Bench Telecom (AA run)32.46Oct 8, 2026aa_tau2
- τ²-Bench Telecom (AA run)54.68Oct 8, 2026aa_tau2
Claude Haiku 4.5: common questions
Who makes Claude Haiku 4.5?
Claude Haiku 4.5 is made by Anthropic.
When was Claude Haiku 4.5 released?
Claude Haiku 4.5 was released on Oct 15, 2025, according to Artificial Analysis.
What is Claude Haiku 4.5 good at?
Claude Haiku 4.5 is capable in long context; and behind the leaders in multimodal tasks, coding, and instruction following. Too few results yet to rate reasoning, agentic tasks, safety, math, multilingual tasks, or factuality.
How much does Claude Haiku 4.5 cost?
Claude Haiku 4.5 costs $1.00 per million input tokens and $5.00 per million output tokens, according to Anthropic's own price page. We track its price at 4 providers. At a mix of three input tokens to one output token, it costs more than 71% of the 331 priced models we track.
How many benchmarks has Claude Haiku 4.5 been tested on?
We track 81 results for Claude Haiku 4.5 on 57 benchmarks from 22 sources, 27 of them independently verified. The latest was recorded on Oct 8, 2026.
Which API features does Claude Haiku 4.5 support?
OpenRouter lists tool calling, structured outputs, and reasoning for Claude Haiku 4.5.
About this record
Where Claude Haiku 4.5's numbers come from, and every name it appears under.
- Tracked since
- Apr 27, 2026
- Newest source mention
- Sep 7, 2026
Where the results come from
Verification: 81 scores · 27 independently verified · 34 aggregator-attributed · 1 vendor cross-reference · 19 vendor-reported. How these tiers are assigned
From 22 sources on 12 sites. Artificial Analysis supplies 34 of them; the 27 independently verified results come from 7 sites. Bars are coloured by trust tier.
- artificialanalysis.ai34
- api.llm-stats.com9
- arcprize.org9
- storage.googleapis.com6
- www-cdn.anthropic.com6
- anthropic.com4
- raw.githubusercontent.com4
- labs.scale.com3
- epoch.ai2
- swebench.com2
- datasets-server.huggingface.co1
- mistral.ai1