Qwen3.5 122B A10B
Qwen3.5 122B A10B is strong in multimodal tasks; capable in long context and instruction following; and behind the leaders in coding, reasoning, and agentic tasks. Too few results yet to rate safety, math, multilingual tasks, or factuality.
Capability profile
Bars show the model's median result as a share of the leading model's, per capability. Select a row to see its results.
Too few results yet to rate Safety, Math, Multilingual or Factuality.
Price
$0.40input$3.20outputper million tokens
From Alibaba's own price page · 4 providers tracked · All prices
Evidence
98results on74benchmarks
- 7 independently verified
- 37 aggregator
- 54 vendor-reported
From 6 sources · latest Oct 8, 2026 · How verification works
API features
Tool callingStructured outputsReasoning
As listed by OpenRouter
Qwen3.5 122B A10B benchmark results
98 results on 74 benchmarks, grouped by capability. Each bar is the result as a share of the capability leader's; every result links to its source.
9.3% behind the leader5 of 6 ranked benchmarks measured
- 92.10Oct 7, 2026
- 83.90Oct 7, 2026
- MathVista87.40Oct 7, 2026MathVista-Mini
- MMMU-Pro74.97Oct 8, 2026aa_mmmu_pro
- CharXiv (reasoning)77.20Oct 7, 2026CharXiv-R
Show 5 more multimodal resultsHide 5 multimodal results
- 1244.87Aug 25, 2026
- 1223May 23, 2026
- MMMU-Pro70.29Oct 8, 2026aa_mmmu_pro
- 76.90Oct 7, 2026
- 82.90Oct 7, 2026
15.4% behind the leader2 of 3 ranked benchmarks measured
- 60.20Oct 7, 2026
- 76.33Oct 8, 2026
19.3% behind the leader2 of 3 ranked benchmarks measured
- IFBench75.71Oct 8, 2026aa_ifbench
- 61.50Oct 7, 2026
33.1% behind the leader6 of 10 ranked benchmarks measured
- 78.90Oct 7, 2026
- 72.00Oct 7, 2026
- SciCode39.70Oct 8, 2026aa_scicode
- Terminal-Bench 2.147.57Oct 8, 2026terminalbenchV21
- 1360.03Aug 29, 2026
- Terminal-Bench Hard31.06Oct 8, 2026aa_terminalbench_hard
Show 3 more coding resultsHide 3 coding results
- SciCode35.65Sep 4, 2026aa_scicode
- Terminal-Bench 2.147.19Oct 8, 2026terminalbenchV21
- Terminal-Bench Hard29.55Oct 8, 2026aa_terminalbench_hard
33.8% behind the leader3 of 6 ranked benchmarks measured
- GPQA Diamond85.66Oct 8, 2026gpqa
- Humanity's Last Exam25.21Oct 8, 2026aa_hle
- 0.57Oct 8, 2026
Show 5 more reasoning resultsHide 5 reasoning results
- 0.86Oct 8, 2026
- GPQA Diamond82.73Oct 8, 2026gpqa
- GPQA Diamond86.60Oct 7, 2026GPQA
- Humanity's Last Exam15.94Oct 8, 2026aa_hle
- 47.50Oct 7, 2026
44.8% behind the leader5 of 7 ranked benchmarks measured
- 63.80Oct 7, 2026
- 58.00Oct 7, 2026
- τ-Bench V3 · Banking15.26Oct 8, 2026tauBanking
- 15.85Oct 8, 2026
- Terminal-Bench 4.00.00Oct 8, 2026
Show 2 more agentic resultsHide 2 agentic results
- 10.63Oct 8, 2026
- τ-Bench V3 · Banking10.31Oct 8, 2026tauBanking
0 of 4 ranked benchmarks measured
- 11.20Aug 29, 2026
- AA-Omniscience · Accuracy19.12Oct 8, 2026omniscienceAccuracy
- AA-Omniscience · Accuracy24.37Oct 8, 2026omniscienceAccuracy
Show 2 more factuality resultsHide 2 factuality results
- AA-Omniscience · Non-hallucination8.41Oct 8, 2026omniscienceNonHallucination
- AA-Omniscience · Non-hallucination12.91Oct 8, 2026omniscienceNonHallucination
More results
Benchmarks outside the capability baskets. They are not ranked against a leader.
- vectara_avg_summary_length86.40Aug 29, 2026Average Summary Length (Words)
- vectara_answer_rate99.80Aug 29, 2026Answer Rate
- vectara_factual_consistency88.80Aug 29, 2026Factual Consistency Rate
- AA Intelligence15.57Oct 8, 2026aa_intelligence_index
- -41.50Oct 8, 2026
- τ²-Bench Telecom (AA run)84.50Oct 8, 2026aa_tau2
Show 45 more resultsHide 45 results
- 9.63Sep 9, 2026
- 16.25Sep 4, 2026
- AA Intelligence17.75Oct 8, 2026aa_intelligence_index
- -54.97Oct 8, 2026
- 93.30Oct 7, 2026
- Artificial Analysis Coding Index45.71Sep 9, 2026aa_coding_index
- Artificial Analysis Coding Index43.34Sep 9, 2026aa_coding_index
- 40.20Oct 7, 2026
- 72.20Oct 7, 2026
- 69.90Oct 7, 2026
- 91.90Oct 7, 2026
- 81.80Oct 7, 2026
- 24.10Oct 7, 2026
- 85.90Oct 7, 2026
- 83.90Oct 7, 2026
- 62.00Oct 7, 2026
- 67.60Oct 7, 2026
- 90.30Oct 7, 2026
- 93.40Aug 31, 2026
- 82.80Oct 7, 2026
- 74.40Oct 7, 2026
- 86.20Oct 7, 2026
- 87.30Oct 7, 2026
- 59.00Oct 7, 2026
- 86.70Oct 7, 2026
- 82.20Oct 7, 2026
- 94.00Oct 7, 2026
- 86.70Oct 7, 2026
- 74.70Oct 7, 2026
- 76.60Oct 7, 2026
- 39.50Oct 7, 2026
- 89.80Oct 7, 2026
- 85.10Oct 7, 2026
- 69.30Oct 7, 2026
- ScreenSpot-Pro (No tools)70.40Oct 7, 2026ScreenSpot Pro
- 44.10Oct 7, 2026
- 61.70Oct 7, 2026
- 67.10Oct 7, 2026
- 79.50Oct 7, 2026
- 49.40Oct 7, 2026
- 82.00Oct 7, 2026
- 33.60Oct 7, 2026
- 60.50Oct 7, 2026
- 9.00Oct 7, 2026
- τ²-Bench Telecom (AA run)93.57Oct 8, 2026aa_tau2
Qwen3.5 122B A10B: common questions
Who makes Qwen3.5 122B A10B?
Qwen3.5 122B A10B is made by Alibaba.
When was Qwen3.5 122B A10B released?
Qwen3.5 122B A10B was released on Feb 24, 2026, according to Artificial Analysis.
What is Qwen3.5 122B A10B good at?
Qwen3.5 122B A10B is strong in multimodal tasks; capable in long context and instruction following; and behind the leaders in coding, reasoning, and agentic tasks. Too few results yet to rate safety, math, multilingual tasks, or factuality.
How much does Qwen3.5 122B A10B cost?
Qwen3.5 122B A10B costs $0.40 per million input tokens and $3.20 per million output tokens, according to Alibaba's own price page. We track its price at 4 providers. At a mix of three input tokens to one output token, it costs more than 60% of the 331 priced models we track.
How many benchmarks has Qwen3.5 122B A10B been tested on?
We track 98 results for Qwen3.5 122B A10B on 74 benchmarks from 6 sources, 7 of them independently verified. The latest was recorded on Oct 8, 2026.
Which API features does Qwen3.5 122B A10B support?
OpenRouter lists tool calling, structured outputs, and reasoning for Qwen3.5 122B A10B.
About this record
Where Qwen3.5 122B A10B's numbers come from, and every name it appears under.
- Tracked since
- Apr 25, 2026
- Newest source mention
- May 2, 2026
Where the results come from
Verification: 98 scores · 7 independently verified · 37 aggregator-attributed · 54 vendor-reported. How these tiers are assigned
From 6 sources on 5 sites. api.llm-stats.com supplies 54 of them; the 7 independently verified results come from 3 sites. Bars are coloured by trust tier.
- api.llm-stats.com54
- artificialanalysis.ai37
- raw.githubusercontent.com4
- datasets-server.huggingface.co2
- lmarena.ai1