Qwen3.5 35B A3B
Qwen3.5 35B A3B is capable in multimodal tasks, long context, and instruction following; and behind the leaders in reasoning and coding. Too few results yet to rate agentic tasks, safety, math, multilingual tasks, or factuality.
Capability profile
Bars show the model's median result as a share of the leading model's, per capability. Select a row to see its results.
Too few results yet to rate Agentic, Safety, Math, Multilingual or Factuality.
Price
$0.25input$2.00outputper million tokens
From Alibaba's own price page · 5 providers tracked · All prices
Evidence
148results on117benchmarks
- 12 independently verified
- 35 aggregator
- 85 vendor-reported
- 16 cross-referenced
From 12 sources · latest Oct 8, 2026 · How verification works
API features
Tool callingStructured outputsReasoning
As listed by OpenRouter
Qwen3.5 35B A3B benchmark results
148 results on 117 benchmarks, grouped by capability. Each bar is the result as a share of the capability leader's; every result links to its source.
13.5% behind the leader5 of 6 ranked benchmarks measured
- 91.00Oct 7, 2026
- MathVista86.20Oct 7, 2026MathVista-Mini
- 81.40Oct 7, 2026
- CharXiv (reasoning)77.50Oct 7, 2026CharXiv-R
- MMMU-Pro72.66Oct 8, 2026aa_mmmu_pro
Show 4 more multimodal resultsHide 4 multimodal results
- MMMU-Pro69.19Oct 8, 2026aa_mmmu_pro
- 75.10Oct 7, 2026
- 81.90Oct 7, 2026
- 65.30Jul 2, 2026
21.7% behind the leader2 of 3 ranked benchmarks measured
- 59.00Oct 7, 2026
- 72.00Oct 8, 2026
23.8% behind the leader2 of 3 ranked benchmarks measured
- IFBench72.52Oct 8, 2026aa_ifbench
- 60.00Oct 7, 2026
34.6% behind the leader3 of 6 ranked benchmarks measured
- GPQA Diamond84.55Oct 8, 2026gpqa
- Humanity's Last Exam21.04Oct 8, 2026aa_hle
- 0.86Oct 8, 2026
Show 6 more reasoning resultsHide 6 reasoning results
- 0.57Oct 8, 2026
- GPQA Diamond81.92Oct 8, 2026gpqa
- GPQA Diamond84.20Oct 7, 2026GPQA
- Humanity's Last Exam13.44Oct 8, 2026aa_hle
- 47.40Oct 7, 2026
- 22.40Jun 6, 2026
41.7% behind the leader8 of 10 ranked benchmarks measured
- 74.60Oct 7, 2026
- 69.20Oct 7, 2026
- 60.30May 3, 2026
- SciCode37.73Sep 4, 2026aa_scicode
- 44.60May 3, 2026
- Terminal-Bench 2.140.82Oct 8, 2026terminalbenchV21
- Terminal-Bench Hard26.52Oct 8, 2026aa_terminalbench_hard
- 1249.71May 22, 2026
Show 4 more coding resultsHide 4 coding results
- 1249.35Jun 19, 2026
- SciCode29.28Sep 4, 2026aa_scicode
- 70.00May 3, 2026
- Terminal-Bench Hard10.61Oct 8, 2026aa_terminalbench_hard
0 of 7 ranked benchmarks measured
- 5.58Oct 8, 2026
- τ-Bench V3 · Banking4.95Oct 8, 2026tauBanking
- 21.52Oct 8, 2026
Show 4 more agentic resultsHide 4 agentic results
- 61.00Oct 7, 2026
- 20.27Jun 15, 2026
- 62.40Jun 6, 2026
- 54.50Oct 7, 2026
0 of 5 ranked benchmarks measured
- 81.82Sep 2, 2026
- 93.33Sep 2, 2026
- 81.82May 17, 2026
Show 4 more math resultsHide 4 math results
- 93.33May 4, 2026
- AIME 202691.00Jun 6, 2026AIME26
- HMMT Feb 202678.70Jun 6, 2026HMMT Feb 26
- 76.80Jun 6, 2026
0 of 4 ranked benchmarks measured
- Vectara HHEM hallucination ratelower is better10.50May 4, 2026
- AA-Omniscience · Accuracy15.72Oct 8, 2026omniscienceAccuracy
- AA-Omniscience · Non-hallucination6.58Oct 8, 2026omniscienceNonHallucination
Show 2 more factuality resultsHide 2 factuality results
- AA-Omniscience · Accuracy20.13Oct 8, 2026omniscienceAccuracy
- AA-Omniscience · Non-hallucination14.57Oct 8, 2026omniscienceNonHallucination
More results
Benchmarks outside the capability baskets. They are not ranked against a leader.
- 61.70Sep 11, 2026
- 94.90May 4, 2026
- 99.80May 4, 2026
- 89.50May 4, 2026
- τ²-Bench Telecom (AA run)86.26Oct 8, 2026aa_tau2
- AA Intelligence15.12Oct 8, 2026aa_intelligence_index
Show 85 more resultsHide 85 results
- 11.82Sep 4, 2026
- 44.11Jun 18, 2026
- AA Intelligence19.33Oct 8, 2026aa_intelligence_index
- -63.02Oct 8, 2026
- -48.10Oct 8, 2026
- 53.18Jun 24, 2026
- 57.87Jun 24, 2026
- 56.27Jun 24, 2026
- 47.73Jun 24, 2026
- 25.98Jun 24, 2026
- 47.58Jun 24, 2026
- 46.13Jun 24, 2026
- 47.10Jun 24, 2026
- AGIEval-En70.51Aug 11, 2026AGIEval-EN (CoT)
- 92.60Oct 7, 2026
- ARC-Challenge95.39Aug 11, 2026ARC-Challenge (25-shot)
- Artificial Analysis Coding Index36.98Sep 9, 2026aa_coding_index
- 30.25Jun 18, 2026
- 38.40Oct 7, 2026
- 67.30Oct 7, 2026
- 69.50Oct 7, 2026
- 90.20Oct 7, 2026
- 80.70Oct 7, 2026
- Claw Eval (pass@3)51.00Aug 24, 2026Claw-Eval Pass^3
- 65.40May 3, 2026
- 22.80Oct 7, 2026
- 85.00Oct 7, 2026
- 83.10Oct 7, 2026
- 64.80Oct 7, 2026
- 80.94Aug 24, 2026
- 84.20Jun 6, 2026
- GSM8K90.07Aug 11, 2026GSM8K (8-shot, CoT)
- 67.90Oct 7, 2026
- 85.61Aug 11, 2026
- HMMT 202589.20Oct 7, 2026HMMT25
- HMMT Feb. 202589.00Jun 6, 2026HMMT Feb 25
- HMMT Nov. 202589.20Jun 6, 2026HMMT Nov 25
- 66.46Aug 11, 2026
- 91.90Aug 31, 2026
- 79.70Oct 7, 2026
- 71.40Oct 7, 2026
- 83.90Oct 7, 2026
- MBPP70.76Aug 11, 2026MBPP (3-shot)
- 27.00Jun 6, 2026
- Minerva Math59.66Aug 11, 2026Minerva Math (4-shot)
- 85.60Oct 7, 2026
- 91.50Aug 24, 2026
- 59.50Oct 7, 2026
- 81.07Aug 11, 2026
- 85.30Oct 7, 2026
- MMLU-Pro64.49Aug 11, 2026MMLU-Pro (5-shot)
- 81.00Oct 7, 2026
- 93.30Oct 7, 2026
- 85.20Oct 7, 2026
- 72.30Oct 7, 2026
- 74.80Oct 7, 2026
- 20.50Jun 6, 2026
- 36.00Oct 7, 2026
- 89.30Oct 7, 2026
- 44.20Aug 11, 2026
- 82.54Aug 11, 2026
- PIQA (Acc.)82.32Aug 11, 2026PIQA (acc)
- 47.70Jun 6, 2026
- 978.00Jun 6, 2026
- 84.10Oct 7, 2026
- 63.50Oct 7, 2026
- 56.43Aug 11, 2026
- 82.36Aug 11, 2026
- ScreenSpot-Pro (No tools)68.60Oct 7, 2026ScreenSpot Pro
- 41.40Oct 7, 2026
- 58.30Oct 7, 2026
- 4.40May 3, 2026
- 63.40Oct 7, 2026
- 81.20Oct 7, 2026
- 40.50Oct 7, 2026
- 28.70Jun 6, 2026
- 80.40Oct 7, 2026
- 31.90Oct 7, 2026
- 29.10Jun 6, 2026
- 57.10Oct 7, 2026
- 59.10Jun 6, 2026
- WinoGrande79.24Aug 11, 2026WinoGrande (5-shot)
- 8.00Oct 7, 2026
- τ²-Bench Telecom (AA run)89.18Oct 8, 2026aa_tau2
- τ³-Bench68.90Jun 6, 2026TAU3-Bench
Qwen3.5 35B A3B: common questions
Who makes Qwen3.5 35B A3B?
Qwen3.5 35B A3B is made by Alibaba.
When was Qwen3.5 35B A3B released?
Qwen3.5 35B A3B was released on Feb 24, 2026, according to Artificial Analysis.
What is Qwen3.5 35B A3B good at?
Qwen3.5 35B A3B is capable in multimodal tasks, long context, and instruction following; and behind the leaders in reasoning and coding. Too few results yet to rate agentic tasks, safety, math, multilingual tasks, or factuality.
How much does Qwen3.5 35B A3B cost?
Qwen3.5 35B A3B costs $0.25 per million input tokens and $2.00 per million output tokens, according to Alibaba's own price page. We track its price at 5 providers. At a mix of three input tokens to one output token, it is cheaper than 50% of the 331 priced models we track.
How many benchmarks has Qwen3.5 35B A3B been tested on?
We track 148 results for Qwen3.5 35B A3B on 117 benchmarks from 12 sources, 12 of them independently verified. The latest was recorded on Oct 8, 2026.
Which API features does Qwen3.5 35B A3B support?
OpenRouter lists tool calling, structured outputs, and reasoning for Qwen3.5 35B A3B.
About this record
Where Qwen3.5 35B A3B's numbers come from, and every name it appears under.
- Tracked since
- Apr 25, 2026
- Newest source mention
- Aug 23, 2026
Where the results come from
Verification: 148 scores · 12 independently verified · 35 aggregator-attributed · 16 vendor cross-reference · 85 vendor-reported. How these tiers are assigned
From 12 sources on 7 sites. api.llm-stats.com supplies 54 of them; the 12 independently verified results come from 4 sites. Bars are coloured by trust tier.
- api.llm-stats.com54
- huggingface.co47
- artificialanalysis.ai35
- matharena.ai4
- raw.githubusercontent.com4
- 99franklin.github.io2
- datasets-server.huggingface.co2