Qwen3.8 2.4T A95B
Qwen3.8 2.4T A95B is capable in long context, agentic tasks, reasoning, and factuality; and behind the leaders in coding. Too few results yet to rate safety, math, multimodal tasks, multilingual tasks, or instruction following.
Capability profile
Bars show the model's median result as a share of the leading model's, per capability. Select a row to see its results.
Too few results yet to rate Safety, Math, Multimodal, Multilingual or Instruction Following.
Price
$2.00input$6.00outputper million tokens
From Alibaba's own price page · 6 providers tracked · All prices
Evidence
19results on17benchmarks
- 4 independently verified
- 15 aggregator
From 5 sources · latest Oct 8, 2026 · How verification works
API features
Tool callingStructured outputsReasoning
As listed by OpenRouter
Qwen3.8 2.4T A95B benchmark results
19 results on 17 benchmarks, grouped by capability. Each bar is the result as a share of the capability leader's; every result links to its source.
10.3% behind the leader1 of 3 ranked benchmarks measured
- 80.33Oct 8, 2026
14.1% behind the leader4 of 7 ranked benchmarks measured
- τ-Bench V3 · Banking49.07Oct 8, 2026tauBanking
- 84.50Oct 8, 2026
- 55.63Oct 8, 2026
- Terminal-Bench 4.011.11Oct 8, 2026
14.7% behind the leader4 of 6 ranked benchmarks measured
- GPQA Diamond93.54Oct 8, 2026gpqa
- 62.50Aug 14, 2026
- Humanity's Last Exam42.45Oct 8, 2026aa_hle
- 20.00Oct 8, 2026
20.8% behind the leader2 of 4 ranked benchmarks measured
- AA-Omniscience · Non-hallucination60.82Oct 8, 2026omniscienceNonHallucination
- AA-Omniscience · Accuracy31.25Oct 8, 2026omniscienceAccuracy
30.5% behind the leader2 of 10 ranked benchmarks measured
- Terminal-Bench 2.182.02Oct 8, 2026terminalbenchV21
- SciCode54.05Oct 8, 2026aa_scicode
More results
Benchmarks outside the capability baskets. They are not ranked against a leader.
- AA Intelligence58.00Aug 14, 2026Artificial Analysis Intelligence Index
- AA Intelligence40.00Aug 14, 2026Artificial Analysis Intelligence Index
- AA Intelligence39.89Oct 8, 2026aa_intelligence_index
- 4.32Oct 8, 2026
- Artificial Analysis Coding Index71.89Sep 9, 2026aa_coding_index
- 50.44Sep 9, 2026
Qwen3.8 2.4T A95B: common questions
Who makes Qwen3.8 2.4T A95B?
Qwen3.8 2.4T A95B is made by Alibaba.
When was Qwen3.8 2.4T A95B released?
Qwen3.8 2.4T A95B was released on Aug 12, 2026, according to Artificial Analysis.
What is Qwen3.8 2.4T A95B good at?
Qwen3.8 2.4T A95B is capable in long context, agentic tasks, reasoning, and factuality; and behind the leaders in coding. Too few results yet to rate safety, math, multimodal tasks, multilingual tasks, or instruction following.
How much does Qwen3.8 2.4T A95B cost?
Qwen3.8 2.4T A95B costs $2.00 per million input tokens and $6.00 per million output tokens, according to Alibaba's own price page. We track its price at 6 providers. At a mix of three input tokens to one output token, it costs more than 75% of the 331 priced models we track.
How many benchmarks has Qwen3.8 2.4T A95B been tested on?
We track 19 results for Qwen3.8 2.4T A95B on 17 benchmarks from 5 sources, 4 of them independently verified. The latest was recorded on Oct 8, 2026.
Which API features does Qwen3.8 2.4T A95B support?
OpenRouter lists tool calling, structured outputs, and reasoning for Qwen3.8 2.4T A95B.
About this record
Where Qwen3.8 2.4T A95B's numbers come from, and every name it appears under.
- Tracked since
- Aug 12, 2026
- Newest source mention
- Oct 7, 2026
Where the results come from
Verification: 19 scores · 4 independently verified · 15 aggregator-attributed. How these tiers are assigned
From 5 sources on 3 sites. Artificial Analysis supplies 17 of them; the 4 independently verified results come from 3 sites. Bars are coloured by trust tier.
- artificialanalysis.ai17
- labs.scale.com1
- simple-bench.com1