Qwen3.5 9B
Qwen3.5 9B is behind the leaders in long context, instruction following, multimodal tasks, coding, and reasoning. Too few results yet to rate agentic tasks, safety, math, multilingual tasks, or factuality.
Capability profile
Bars show the model's median result as a share of the leading model's, per capability. Select a row to see its results.
Too few results yet to rate Agentic, Safety, Math, Multilingual or Factuality.
Price
$0.14input$0.20outputper million tokens
From Artificial Analysis · 5 providers tracked · All prices
Evidence
91results on57benchmarks
- 5 independently verified
- 37 aggregator
- 27 vendor-reported
- 22 cross-referenced
From 11 sources · latest Oct 8, 2026 · How verification works
API features
Tool callingStructured outputsReasoning
As listed by OpenRouter
Research
83 papers reference Qwen3.5 9BQwen3.5 9B benchmark results
91 results on 57 benchmarks, grouped by capability. Each bar is the result as a share of the capability leader's; every result links to its source.
27.0% behind the leader2 of 3 ranked benchmarks measured
- 55.20Oct 7, 2026
- 70.00Oct 8, 2026
31.4% behind the leader2 of 3 ranked benchmarks measured
- IFBench66.73Oct 8, 2026aa_ifbench
- 54.50Oct 7, 2026
32.4% behind the leader1 of 6 ranked benchmarks measured
- MMMU-Pro69.25Oct 8, 2026aa_mmmu_pro
Show 2 more multimodal resultsHide 2 multimodal results
- MMMU-Pro66.76Oct 8, 2026aa_mmmu_pro
- 58.70May 22, 2026
39.8% behind the leader4 of 10 ranked benchmarks measured
- 65.60Oct 7, 2026
- SciCode29.51Oct 8, 2026aa_scicode
- Terminal-Bench Hard24.24Oct 8, 2026aa_terminalbench_hard
- Terminal-Bench 2.129.21Oct 8, 2026terminalbenchV21
Show 4 more coding resultsHide 4 coding results
- SciCode27.68Sep 4, 2026aa_scicode
- Terminal-Bench 2.121.35Oct 8, 2026terminalbenchV21
- 27.00Sep 24, 2026
- Terminal-Bench Hard18.18Oct 8, 2026aa_terminalbench_hard
40.0% behind the leader3 of 6 ranked benchmarks measured
- GPQA Diamond80.61Oct 8, 2026gpqa
- Humanity's Last Exam14.92Oct 8, 2026aa_hle
- 0.29Oct 8, 2026
Show 4 more reasoning resultsHide 4 reasoning results
- 0.57Oct 8, 2026
- GPQA Diamond78.59Oct 8, 2026gpqa
- GPQA Diamond81.70Oct 7, 2026GPQA
- Humanity's Last Exam9.36Oct 8, 2026aa_hle
0 of 7 ranked benchmarks measured
- 0.00Oct 8, 2026
- Terminal-Bench 4.00.51Oct 8, 2026
- τ-Bench V3 · Banking7.01Oct 8, 2026tauBanking
Show 2 more agentic resultsHide 2 agentic results
- 17.16Jun 15, 2026
- τ-Bench V3 · Banking4.12Aug 10, 2026tauBanking
0 of 5 ranked benchmarks measured
- 71.21Sep 2, 2026
- 92.50Sep 2, 2026
- 71.21May 10, 2026
0 of 4 ranked benchmarks measured
- AA-Omniscience · Accuracy13.98Oct 8, 2026omniscienceAccuracy
- AA-Omniscience · Accuracy16.42Oct 8, 2026omniscienceAccuracy
- AA-Omniscience · Non-hallucination1.55Oct 8, 2026omniscienceNonHallucination
Show 1 more factuality resultHide 1 factuality result
- AA-Omniscience · Non-hallucination16.39Oct 8, 2026omniscienceNonHallucination
More results
Benchmarks outside the capability baskets. They are not ranked against a leader.
- 64.10Sep 11, 2026
- AA Intelligence13.27Oct 8, 2026aa_intelligence_index
- τ²-Bench Telecom (AA run)85.09Oct 8, 2026aa_tau2
- -70.70Oct 8, 2026
- τ²-Bench Telecom (AA run)86.84Oct 8, 2026aa_tau2
- AA Intelligence11.20Oct 8, 2026aa_intelligence_index
Show 45 more resultsHide 45 results
- 7.00Sep 4, 2026
- 41.10Jun 18, 2026
- AA Intelligence21.00Jul 4, 2026Artificial Analysis Intelligence Index
- -53.47Oct 8, 2026
- AIME25 no tools56.07Aug 10, 2026AIME25
- AIME25 no tools56.07Aug 9, 2026AIME25
- Artificial Analysis Coding Index28.66Sep 9, 2026aa_coding_index
- Artificial Analysis Coding Index23.46Sep 9, 2026aa_coding_index
- 66.10Oct 7, 2026
- 60.13Aug 10, 2026
- 60.13Aug 9, 2026
- 27.23Aug 10, 2026
- 27.23Aug 9, 2026
- 88.20Oct 7, 2026
- Claw-Eval Avg66.53Aug 10, 2026Claw-Eval average (EN)
- Claw-Eval Avg66.53Aug 9, 2026Claw-Eval average (EN)
- 18.00Oct 7, 2026
- 82.90Oct 7, 2026
- 91.50Aug 31, 2026
- 75.60Oct 7, 2026
- 2.60Sep 24, 2026
- LiveCodeBenchV6 no tools69.86Aug 10, 2026LiveCodeBenchv6
- LiveCodeBenchV6 no tools69.86Aug 9, 2026LiveCodeBenchv6
- 52.39Jun 15, 2026
- 82.50Oct 7, 2026
- 76.30Oct 7, 2026
- 91.10Oct 7, 2026
- 81.20Oct 7, 2026
- 68.64Jun 15, 2026
- 91.78Jun 15, 2026
- 9.00Sep 24, 2026
- 71.45Aug 10, 2026
- 71.45Aug 9, 2026
- 93.62Jun 15, 2026
- 88.83Jun 15, 2026
- 58.20Oct 7, 2026
- 32.00Sep 24, 2026
- SWE Verified60.00Sep 24, 2026
- 79.10Oct 7, 2026
- 25.90Sep 24, 2026
- 56.00Jun 15, 2026
- 29.80Oct 7, 2026
- 83.22Jun 15, 2026
- 5.15Aug 10, 2026
- 5.15Aug 9, 2026
Qwen3.5 9B: common questions
Who makes Qwen3.5 9B?
Qwen3.5 9B is made by Alibaba.
When was Qwen3.5 9B released?
Qwen3.5 9B was released on Mar 2, 2026, according to Artificial Analysis.
What is Qwen3.5 9B good at?
Qwen3.5 9B is behind the leaders in long context, instruction following, multimodal tasks, coding, and reasoning. Too few results yet to rate agentic tasks, safety, math, multilingual tasks, or factuality.
How much does Qwen3.5 9B cost?
Qwen3.5 9B costs $0.14 per million input tokens and $0.20 per million output tokens, according to Artificial Analysis. We track its price at 5 providers. At a mix of three input tokens to one output token, it is cheaper than 84% of the 330 priced models we track.
How many benchmarks has Qwen3.5 9B been tested on?
We track 91 results for Qwen3.5 9B on 57 benchmarks from 11 sources, 5 of them independently verified. The latest was recorded on Oct 8, 2026.
Which API features does Qwen3.5 9B support?
OpenRouter lists tool calling, structured outputs, and reasoning for Qwen3.5 9B.
About this record
Where Qwen3.5 9B's numbers come from, and every name it appears under.
- Tracked since
- Apr 27, 2026
- Newest source mention
- Aug 29, 2026
Where the results come from
Verification: 91 scores · 5 independently verified · 37 aggregator-attributed · 22 vendor cross-reference · 27 vendor-reported. How these tiers are assigned
From 11 sources on 5 sites. Artificial Analysis supplies 38 of them; the 5 independently verified results come from 2 sites. Bars are coloured by trust tier.
- artificialanalysis.ai38
- huggingface.co29
- api.llm-stats.com19
- matharena.ai3
- 99franklin.github.io2