Qwen3.8 Max Preview
Qwen3.8 Max Preview is at the frontier in long context; strong in agentic tasks; capable in instruction following, factuality, reasoning, coding, and multimodal tasks; and behind the leaders in math. Too few results yet to rate safety or multilingual tasks.
Capability profile
Bars show the model's median result as a share of the leading model's, per capability. Select a row to see its results.
Too few results yet to rate Safety or Multilingual.
Price
$2.00input$6.00outputper million tokens
From Artificial Analysis · 4 providers tracked · All prices
Evidence
87results on59benchmarks
- 19 independently verified
- 32 aggregator
- 27 vendor-reported
- 9 cross-referenced
From 12 sources · latest Oct 8, 2026 · How verification works
API features
Tool callingStructured outputsReasoning
As listed by OpenRouter
Qwen3.8 Max Preview benchmark results
87 results on 59 benchmarks, grouped by capability. Each bar is the result as a share of the capability leader's; every result links to its source.
1.8% behind the leader2 of 3 ranked benchmarks measured
- 66.30Oct 7, 2026
- 80.33Oct 8, 2026
Show 1 more long context resultHide 1 long context result
- 78.33Oct 8, 2026
9.2% behind the leader5 of 7 ranked benchmarks measured
- 86.10Oct 7, 2026
- τ-Bench V3 · Banking47.84Oct 8, 2026tauBanking
- 58.57Oct 8, 2026
- Terminal-Bench 4.038.89Oct 8, 2026
Show 5 more agentic resultsHide 5 agentic results
- 42.40Oct 8, 2026
- 40.25Oct 8, 2026
- 55.38Oct 8, 2026
- Terminal-Bench 4.018.69Oct 8, 2026
- τ-Bench V3 · Banking51.34Oct 8, 2026tauBanking
10.8% behind the leader2 of 3 ranked benchmarks measured
- 82.80Oct 7, 2026
- LiveBench · Instruction Following74.08Oct 8, 2026livebench_instruction_following@2026-06-25
14.0% behind the leader3 of 4 ranked benchmarks measured
- AA-Omniscience · Non-hallucination71.16Oct 8, 2026omniscienceNonHallucination
- 47.30Sep 21, 2026
- AA-Omniscience · Accuracy31.68Oct 8, 2026omniscienceAccuracy
Show 3 more factuality resultsHide 3 factuality results
- AA-Omniscience · Accuracy31.85Oct 8, 2026omniscienceAccuracy
- AA-Omniscience · Non-hallucination58.25Oct 8, 2026omniscienceNonHallucination
- 45.80Sep 21, 2026
15.4% behind the leader4 of 6 ranked benchmarks measured
- GPQA Diamond92.83Oct 8, 2026gpqa
- LiveBench · Reasoning88.21Oct 8, 2026livebench_reasoning@2026-06-25
- Humanity's Last Exam43.10Oct 8, 2026aa_hle
- 17.71Oct 8, 2026
Show 5 more reasoning resultsHide 5 reasoning results
- 20.00Oct 8, 2026
- GPQA Diamond92.73Oct 8, 2026gpqa
- GPQA Diamond92.60Oct 7, 2026GPQA
- Humanity's Last Exam43.05Oct 8, 2026aa_hle
- 43.60Oct 7, 2026
15.6% behind the leader6 of 10 ranked benchmarks measured
- Terminal-Bench 2.188.76Oct 8, 2026terminalbenchV21
- 1672.04Sep 20, 2026
- LiveBench · Agentic Coding64.65Oct 8, 2026livebench_agentic_coding@2026-06-25
- LiveBench · Coding72.87Oct 8, 2026livebench_coding@2026-06-25
- SciCode52.08Oct 8, 2026aa_scicode
- 67.70Oct 7, 2026
Show 5 more coding resultsHide 5 coding results
- 1672.20Sep 20, 2026
- SciCode53.24Oct 8, 2026aa_scicode
- Terminal-Bench 2.181.27Oct 8, 2026terminalbenchV21
- 86.60Oct 7, 2026
- 86.60Aug 28, 2026
23.1% behind the leader1 of 6 ranked benchmarks measured
- MMMU-Pro82.77Oct 8, 2026aa_mmmu_pro
Show 3 more multimodal resultsHide 3 multimodal results
- 1314.20Sep 20, 2026
- MMMU-Pro82.31Oct 8, 2026aa_mmmu_pro
- 82.30Oct 7, 2026
26.0% behind the leader5 of 5 ranked benchmarks measured
- LiveBench · Mathematics91.31Oct 8, 2026livebench_math@2026-06-25
- 74.74Sep 21, 2026
- 46.34Sep 21, 2026
Show 2 more math resultsHide 2 math results
- 34.15Sep 21, 2026
- 65.61Sep 21, 2026
More results
Benchmarks outside the capability baskets. They are not ranked against a leader.
- livebench_language79.69Oct 8, 2026livebench_language@2026-06-25
- livebench_data_analysis78.41Oct 8, 2026livebench_data_analysis@2026-06-25
- 1482Oct 6, 2026
- AA Intelligence45.00Sep 20, 2026Artificial Analysis Intelligence Index
- AA Intelligence40.00Sep 15, 2026Artificial Analysis Intelligence Index
- 11.98Oct 8, 2026
Show 32 more resultsHide 32 results
- 49.59Sep 9, 2026
- AA Intelligence45.42Oct 8, 2026aa_intelligence_index
- AA Intelligence40.15Oct 8, 2026aa_intelligence_index
- 3.40Oct 8, 2026
- 52.40Oct 7, 2026
- 27.00Aug 28, 2026
- 52.40Aug 12, 2026
- 75.10Oct 7, 2026
- 85.30Oct 7, 2026
- Artificial Analysis Coding Index71.81Sep 9, 2026aa_coding_index
- 78.50Aug 28, 2026
- DeepSWE56.60Aug 28, 2026DeepSWE (v1.1)
- 56.60Oct 7, 2026
- 77.80Oct 7, 2026
- 1739.00Aug 28, 2026
- 60.20Oct 7, 2026
- HLE (with tools)56.20Aug 12, 2026HLE w/ tools
- HLE (with tools)56.20Aug 28, 2026HLE w/ Tools
- 53.40Oct 7, 2026
- 81.80Oct 7, 2026
- MLS Bench Litelower is better41.00Oct 7, 2026
- 92.90Aug 24, 2026
- 55.90Oct 7, 2026
- 55.90Aug 28, 2026
- 93.00Oct 7, 2026
- 10.50Aug 28, 2026
- 88.00Oct 7, 2026
- ScreenSpot-Pro (No tools)84.50Oct 7, 2026ScreenSpot Pro
- 72.50Oct 7, 2026
- Toolathlon Verified72.50Aug 12, 2026Toolathlon Verified (Pass@1)
- 72.50Aug 28, 2026
- 81.90Oct 7, 2026
Qwen3.8 Max Preview: common questions
Who makes Qwen3.8 Max Preview?
Qwen3.8 Max Preview is made by Alibaba.
When was Qwen3.8 Max Preview released?
Qwen3.8 Max Preview was released on Aug 3, 2026, according to Artificial Analysis.
What is Qwen3.8 Max Preview good at?
Qwen3.8 Max Preview is at the frontier in long context; strong in agentic tasks; capable in instruction following, factuality, reasoning, coding, and multimodal tasks; and behind the leaders in math. Too few results yet to rate safety or multilingual tasks.
How much does Qwen3.8 Max Preview cost?
Qwen3.8 Max Preview costs $2.00 per million input tokens and $6.00 per million output tokens, according to Artificial Analysis. We track its price at 4 providers. At a mix of three input tokens to one output token, it costs more than 75% of the 331 priced models we track.
How many benchmarks has Qwen3.8 Max Preview been tested on?
We track 87 results for Qwen3.8 Max Preview on 59 benchmarks from 12 sources, 19 of them independently verified. The latest was recorded on Oct 8, 2026.
Which API features does Qwen3.8 Max Preview support?
OpenRouter lists tool calling, structured outputs, and reasoning for Qwen3.8 Max Preview.
About this record
Where Qwen3.8 Max Preview's numbers come from, and every name it appears under.
- Tracked since
- Jul 19, 2026
- Newest source mention
- Sep 22, 2026
Where the results come from
Verification: 87 scores · 19 independently verified · 32 aggregator-attributed · 9 vendor cross-reference · 27 vendor-reported. How these tiers are assigned
From 12 sources on 7 sites. Artificial Analysis supplies 34 of them; the 19 independently verified results come from 5 sites. Bars are coloured by trust tier.
- artificialanalysis.ai34
- api.llm-stats.com23
- huggingface.co13
- livebench.ai7
- epoch.ai6
- datasets-server.huggingface.co3
- lmarena.ai1