Qwen3.5 0.8B
Qwen3.5 0.8B is behind the leaders in multimodal tasks. Too few results yet to rate reasoning, coding, agentic tasks, safety, long context, math, multilingual tasks, instruction following, or factuality.
Capability profile
Bars show the model's median result as a share of the leading model's, per capability. Select a row to see its results.
Too few results yet to rate Reasoning, Coding, Agentic, Safety, Long Context, Math, Multilingual, Instruction Following or Factuality.
Price
No current price is tracked for this model. See the rate card
Evidence
67results on43benchmarks
- 33 aggregator
- 26 vendor-reported
- 8 cross-referenced
From 5 sources · latest Oct 8, 2026 · How verification works
Research
12 papers reference Qwen3.5 0.8BQwen3.5 0.8B benchmark results
67 results on 43 benchmarks, grouped by capability. Each bar is the result as a share of the capability leader's; every result links to its source.
44.3% behind the leader1 of 6 ranked benchmarks measured
- MMMU-Pro25.84Oct 8, 2026aa_mmmu_pro
Show 1 more multimodal resultHide 1 multimodal result
- MMMU-Pro25.66Oct 8, 2026aa_mmmu_pro
0 of 6 ranked benchmarks measured
- 0.00Oct 8, 2026
- GPQA Diamond11.11Oct 8, 2026gpqa
- Humanity's Last Exam1.07Oct 8, 2026aa_hle
Show 5 more reasoning resultsHide 5 reasoning results
- GPQA Diamond23.64Oct 8, 2026gpqa
- GPQA Diamond11.90Oct 7, 2026GPQA
- 26.87Aug 15, 2026
- 19.29Aug 9, 2026
- Humanity's Last Exam5.10Oct 8, 2026aa_hle
0 of 10 ranked benchmarks measured
- Terminal-Bench 2.10.00Oct 8, 2026terminalbenchV21
- Terminal-Bench Hard0.00Oct 8, 2026aa_terminalbench_hard
- SciCode0.00Sep 4, 2026aa_scicode
Show 2 more coding resultsHide 2 coding results
- SciCode2.89Sep 4, 2026aa_scicode
- Terminal-Bench 2.10.37Aug 12, 2026terminalbenchV21
0 of 7 ranked benchmarks measured
- 0.00Oct 8, 2026
- τ-Bench V3 · Banking0.00Aug 10, 2026tauBanking
- τ-Bench V3 · Banking1.24Aug 10, 2026tauBanking
0 of 3 ranked benchmarks measured
- 9.00Oct 8, 2026
- 8.00Oct 8, 2026
- 26.10Oct 7, 2026
Show 1 more long context resultHide 1 long context result
- 4.70Oct 7, 2026
0 of 3 ranked benchmarks measured
0 of 4 ranked benchmarks measured
- AA-Omniscience · Accuracy4.23Oct 8, 2026omniscienceAccuracy
- AA-Omniscience · Non-hallucination38.65Oct 8, 2026omniscienceNonHallucination
- AA-Omniscience · Accuracy4.98Oct 8, 2026omniscienceAccuracy
Show 1 more factuality resultHide 1 factuality result
- AA-Omniscience · Non-hallucination1.53Oct 8, 2026omniscienceNonHallucination
More results
Benchmarks outside the capability baskets. They are not ranked against a leader.
- τ²-Bench Telecom (AA run)47.66Oct 8, 2026aa_tau2
- AA Intelligence6.13Oct 8, 2026aa_intelligence_index
- -54.52Oct 8, 2026
- τ²-Bench Telecom (AA run)65.20Oct 8, 2026aa_tau2
- -88.58Oct 8, 2026
- AA Intelligence5.40Oct 8, 2026aa_intelligence_index
Show 30 more resultsHide 30 results
- 5.70Aug 10, 2026
- 0.41Aug 10, 2026
- AIME 20245.67Aug 15, 2026AIME 24
- AIME 20258.67Aug 15, 2026AIME 25
- Artificial Analysis Coding Index0.00Sep 9, 2026aa_coding_index
- 1.21Aug 12, 2026
- 64.02Aug 15, 2026
- 39.64Aug 9, 2026
- 25.30Oct 7, 2026
- 25.39Aug 9, 2026
- 50.50Oct 7, 2026
- 53.30Aug 15, 2026
- 46.81Aug 15, 2026
- 44.00Aug 31, 2026
- 71.23Aug 15, 2026
- 32.93Aug 9, 2026
- 40.60Oct 7, 2026
- MATH-500 (EM)57.60Aug 15, 2026Math-500
- 42.33Aug 15, 2026
- 72.92Aug 15, 2026
- 42.30Oct 7, 2026
- 34.60Oct 7, 2026
- 38.93Aug 15, 2026
- 59.50Oct 7, 2026
- 44.30Oct 7, 2026
- 21.30Oct 7, 2026
- 11.60Oct 7, 2026
- 82.30Aug 9, 2026
- τ²-Bench14.33Aug 9, 2026τ²-Bench Telecom
- τ²-Bench (Retail)7.02Aug 9, 2026τ²-Bench Retail
Qwen3.5 0.8B: common questions
Who makes Qwen3.5 0.8B?
Qwen3.5 0.8B is made by Alibaba.
When was Qwen3.5 0.8B released?
Qwen3.5 0.8B was released on Mar 2, 2026, according to Artificial Analysis.
What is Qwen3.5 0.8B good at?
Qwen3.5 0.8B is behind the leaders in multimodal tasks. Too few results yet to rate reasoning, coding, agentic tasks, safety, long context, math, multilingual tasks, instruction following, or factuality.
How many benchmarks has Qwen3.5 0.8B been tested on?
We track 67 results for Qwen3.5 0.8B on 43 benchmarks from 5 sources. The latest was recorded on Oct 8, 2026.
About this record
Where Qwen3.5 0.8B's numbers come from, and every name it appears under.
- Tracked since
- May 2, 2026
- Newest source mention
- Sep 18, 2026
Where the results come from
Verification: 67 scores · 0 independently verified · 33 aggregator-attributed · 8 vendor cross-reference · 26 vendor-reported. How these tiers are assigned
From 5 sources on 3 sites. Artificial Analysis supplies 33 of them. Bars are coloured by trust tier.
- artificialanalysis.ai33
- huggingface.co19
- api.llm-stats.com15