Qwen3.5 397B A17B
Qwen3.5 397B A17B is strong in long context; capable in instruction following and multimodal tasks; and behind the leaders in coding, reasoning, and agentic tasks. Too few results yet to rate safety, math, multilingual tasks, or factuality.
Capability profile
Bars show the model's median result as a share of the leading model's, per capability. Select a row to see its results.
Too few results yet to rate Safety, Math, Multilingual or Factuality.
Price
$0.60input$3.60outputper million tokens
From Alibaba's own price page · 7 providers tracked · All prices
Evidence
163results on112benchmarks
- 11 independently verified
- 37 aggregator
- 64 vendor-reported
- 51 cross-referenced
From 16 sources · latest Oct 8, 2026 · How verification works
API features
Tool callingStructured outputsReasoning
As listed by OpenRouter
Qwen3.5 397B A17B benchmark results
163 results on 112 benchmarks, grouped by capability. Each bar is the result as a share of the capability leader's; every result links to its source.
9.0% behind the leader2 of 3 ranked benchmarks measured
- 63.20Oct 7, 2026
- 77.33Oct 8, 2026
Show 3 more long context resultsHide 3 long context results
- 64.33Oct 8, 2026
- 68.70Oct 7, 2026
- LongBench v268.90Jun 10, 2026Longbench v2 (≤ 1M)
14.0% behind the leader2 of 3 ranked benchmarks measured
- IFBench78.78Oct 8, 2026aa_ifbench
- 67.60Oct 7, 2026
Show 4 more instruction following resultsHide 4 instruction following results
- IFBench51.63Oct 8, 2026aa_ifbench
- 76.50Oct 7, 2026
- 78.80Aug 24, 2026
- 63.90Jun 10, 2026
19.3% behind the leader3 of 6 ranked benchmarks measured
- 85.00Jun 15, 2026
- MMMU-Pro77.28Oct 8, 2026aa_mmmu_pro
- CharXiv (reasoning)80.80Aug 24, 2026CharXiv RQ
Show 6 more multimodal resultsHide 6 multimodal results
28.8% behind the leader8 of 10 ranked benchmarks measured
- 83.60Oct 7, 2026
- 76.40Oct 7, 2026
- 69.30Oct 7, 2026
- SciCode44.79Oct 8, 2026aa_scicode
- Terminal-Bench Hard40.91Oct 8, 2026aa_terminalbench_hard
- 1399.85May 22, 2026
- 50.90Jun 15, 2026
- Terminal-Bench 2.151.31Oct 8, 2026terminalbenchV21
Show 10 more coding resultsHide 10 coding results
- LiveCodeBench v679.30Jun 6, 2026LiveCodeBench (v6)
- SciCode41.09Sep 4, 2026aa_scicode
- 42.00Jul 30, 2026
- 70.90Jun 6, 2026
- 76.20Jun 15, 2026
- 76.40Jul 30, 2026
- 73.60Jun 6, 2026
- 51.30Jul 30, 2026
- 49.90Jun 6, 2026
- Terminal-Bench Hard35.61Oct 8, 2026aa_terminalbench_hard
29.2% behind the leader3 of 6 ranked benchmarks measured
- GPQA Diamond89.29Oct 8, 2026gpqa
- Humanity's Last Exam28.96Oct 8, 2026aa_hle
- 1.71Oct 8, 2026
Show 10 more reasoning resultsHide 10 reasoning results
- 0.86Oct 8, 2026
- 1.70Jul 30, 2026
- GPQA Diamond86.06Oct 8, 2026gpqa
- GPQA Diamond88.40Oct 7, 2026GPQA
- GPQA Diamond87.10Aug 24, 2026GPQA (no tools)
- 89.30Jul 30, 2026
- Humanity's Last Exam19.83Oct 8, 2026aa_hle
- 28.70Oct 7, 2026
- Humanity's Last Exam27.30Jul 30, 2026HLE text only
- Humanity's Last Exam28.50Jun 6, 2026HLE (no tools)
44.0% behind the leader4 of 7 ranked benchmarks measured
- 69.00Oct 7, 2026
- τ-Bench V3 · Banking13.40Oct 8, 2026tauBanking
- 14.79Oct 8, 2026
- Terminal-Bench 4.00.00Oct 8, 2026
Show 6 more agentic resultsHide 6 agentic results
- 15.34Oct 8, 2026
- 34.09Oct 8, 2026
- 40.50Jun 6, 2026
- 35.83Jun 15, 2026
- GDPval (win rate)34.60Jun 6, 2026GDPVal
- 20.90Jun 6, 2026
0 of 5 ranked benchmarks measured
- 31.23Sep 9, 2026
- 36.31Sep 2, 2026
- 87.88Sep 2, 2026
Show 9 more math resultsHide 9 math results
- 94.17Sep 2, 2026
- 94.17May 2, 2026
- 91.30Oct 7, 2026
- AIME 202693.30Jun 15, 2026AIME26
- 93.30Jul 30, 2026
- 87.88May 10, 2026
- HMMT Feb 202687.90Jun 15, 2026HMMT Feb 26
- 87.90Jul 30, 2026
- 80.90Oct 7, 2026
0 of 4 ranked benchmarks measured
- AA-Omniscience · Accuracy24.52Oct 8, 2026omniscienceAccuracy
- AA-Omniscience · Accuracy30.80Oct 8, 2026omniscienceAccuracy
- AA-Omniscience · Non-hallucination17.31Oct 8, 2026omniscienceNonHallucination
Show 2 more factuality resultsHide 2 factuality results
- AA-Omniscience · Non-hallucination11.05Oct 8, 2026omniscienceNonHallucination
- 26.00Aug 24, 2026
More results
Benchmarks outside the capability baskets. They are not ranked against a leader.
- 1442Jul 23, 2026
- AA Intelligence21.00Jun 20, 2026Artificial Analysis Intelligence Index
- τ²-Bench Telecom (AA run)83.92Oct 8, 2026aa_tau2
- AA Intelligence21.44Oct 8, 2026aa_intelligence_index
- -37.90Oct 8, 2026
- AA Intelligence18.42Oct 8, 2026aa_intelligence_index
Show 79 more resultsHide 79 results
- 10.56Sep 9, 2026
- 53.32Jun 18, 2026
- -30.75Oct 8, 2026
- 54.90Jun 24, 2026
- 68.31Jun 24, 2026
- 60.85Jun 24, 2026
- 54.74Jun 24, 2026
- 30.81Jun 24, 2026
- 64.44Jun 24, 2026
- 55.30Jun 24, 2026
- 48.55Jun 24, 2026
- 61.40Jun 6, 2026
- 60.40Jun 6, 2026
- Artificial Analysis Coding Index48.21Sep 9, 2026aa_coding_index
- 37.43Jun 18, 2026
- 72.90Oct 7, 2026
- BrowseComp (context management)78.60Aug 24, 2026BrowseComp (with context management)
- 70.30Oct 7, 2026
- 93.00Oct 7, 2026
- 82.00Aug 24, 2026
- Claw Eval (pass@3)48.10Aug 24, 2026Claw-Eval Pass^3
- 70.70Jun 15, 2026
- 2.40Jun 6, 2026
- 34.30Oct 7, 2026
- 86.30Aug 24, 2026
- 67.50Aug 24, 2026
- 25.40Jul 8, 2026
- FLTEval pass@425.40Oct 6, 2026
- 962.00Jul 30, 2026
- 90.00Aug 24, 2026
- GPQA (unspecified)87.10Jun 6, 2026GPQA (no tools)
- HLE (with tools)48.30Jul 30, 2026HLE with tools
- 92.70Oct 7, 2026
- HMMT Feb. 202594.80Jun 15, 2026HMMT Feb 25
- HMMT Nov. 202592.70Jun 15, 2026HMMT Nov 25
- 78.20Jun 10, 2026
- 92.60Aug 31, 2026
- 84.51Jun 6, 2026
- 85.60Oct 7, 2026
- 441.30Jun 6, 2026
- 46.10Oct 7, 2026
- 86.70Aug 24, 2026
- 87.80Oct 7, 2026
- 88.30Jun 6, 2026
- 84.70Oct 7, 2026
- 86.40Jun 10, 2026
- 94.90Oct 7, 2026
- 88.50Oct 7, 2026
- 77.60Aug 24, 2026
- 32.20Jun 15, 2026
- 86.60Jun 6, 2026
- 53.00Jun 6, 2026
- 51.80Jun 15, 2026
- 1186.00Jun 15, 2026
- 83.90Aug 24, 2026
- 48.00Jun 6, 2026
- 46.90Oct 7, 2026
- 67.10Aug 24, 2026
- 30.00Jun 15, 2026
- 99.40Aug 24, 2026
- 70.40Oct 7, 2026
- 50.90Jul 30, 2026
- 86.70Oct 7, 2026
- 76.50Jun 6, 2026
- 71.00Jun 6, 2026
- 88.50Jun 6, 2026
- TauBench V3 - Telecom98.00Aug 24, 2026Telecom
- 52.50Oct 7, 2026
- 38.30Oct 7, 2026
- 40.70Jul 30, 2026
- 59.00Jun 6, 2026
- 61.30Jun 6, 2026
- 87.50Aug 24, 2026
- 84.70Aug 24, 2026
- 49.70Oct 7, 2026
- 74.00Oct 7, 2026
- 86.80Jun 10, 2026
- τ²-Bench Telecom (AA run)95.61Oct 8, 2026aa_tau2
- τ³-Bench Banking13.40Jul 30, 2026τ³-Banking
Qwen3.5 397B A17B: common questions
Who makes Qwen3.5 397B A17B?
Qwen3.5 397B A17B is made by Alibaba.
When was Qwen3.5 397B A17B released?
Qwen3.5 397B A17B was released on Feb 16, 2026, according to Artificial Analysis.
What is Qwen3.5 397B A17B good at?
Qwen3.5 397B A17B is strong in long context; capable in instruction following and multimodal tasks; and behind the leaders in coding, reasoning, and agentic tasks. Too few results yet to rate safety, math, multilingual tasks, or factuality.
How much does Qwen3.5 397B A17B cost?
Qwen3.5 397B A17B costs $0.60 per million input tokens and $3.60 per million output tokens, according to Alibaba's own price page. We track its price at 7 providers. At a mix of three input tokens to one output token, it costs more than 64% of the 330 priced models we track.
How many benchmarks has Qwen3.5 397B A17B been tested on?
We track 163 results for Qwen3.5 397B A17B on 112 benchmarks from 16 sources, 11 of them independently verified. The latest was recorded on Oct 8, 2026.
Which API features does Qwen3.5 397B A17B support?
OpenRouter lists tool calling, structured outputs, and reasoning for Qwen3.5 397B A17B.
About this record
Where Qwen3.5 397B A17B's numbers come from, and every name it appears under.
- Tracked since
- Apr 27, 2026
- Newest source mention
- Aug 23, 2026
Where the results come from
Verification: 163 scores · 11 independently verified · 37 aggregator-attributed · 51 vendor cross-reference · 64 vendor-reported. How these tiers are assigned
From 16 sources on 9 sites. Hugging Face supplies 69 of them; the 11 independently verified results come from 5 sites. Bars are coloured by trust tier.
- huggingface.co69
- artificialanalysis.ai38
- api.llm-stats.com31
- thinkingmachines.ai13
- matharena.ai5
- datasets-server.huggingface.co2
- lmarena.ai2
- mistral.ai2
- epoch.ai1