Qwen3 VL 32B Reasoning
Qwen3 VL 32B Reasoning is behind the leaders in multimodal tasks, instruction following, long context, and agentic tasks. Too few results yet to rate reasoning, coding, safety, math, multilingual tasks, or factuality.
Capability profile
Bars show the model's median result as a share of the leading model's, per capability. Select a row to see its results.
Too few results yet to rate Reasoning, Coding, Safety, Math, Multilingual or Factuality.
Price
$0.16input$0.64outputper million tokens
From Alibaba's own price page · 2 providers tracked · All prices
Evidence
52results on50benchmarks
- 16 aggregator
- 36 vendor-reported
From 2 sources · latest Oct 8, 2026 · How verification works
Qwen3 VL 32B Reasoning benchmark results
52 results on 50 benchmarks, grouped by capability. Each bar is the result as a share of the capability leader's; every result links to its source.
32.0% behind the leader4 of 6 ranked benchmarks measured
- MathVista85.90Oct 7, 2026MathVista-Mini
- 85.50Oct 7, 2026
- MMMU-Pro63.41Oct 8, 2026aa_mmmu_pro
- CharXiv (reasoning)65.20Oct 7, 2026CharXiv-R
34.0% behind the leader1 of 3 ranked benchmarks measured
- IFBench59.39Oct 8, 2026aa_ifbench
34.0% behind the leader1 of 3 ranked benchmarks measured
- 55.33Oct 8, 2026
45.5% behind the leader2 of 7 ranked benchmarks measured
- OSWorld-Verified41.00Oct 7, 2026OSWorld
- 7.30Jun 15, 2026
0 of 6 ranked benchmarks measured
- 0.00Oct 8, 2026
- GPQA Diamond73.33Oct 8, 2026gpqa
- Humanity's Last Exam10.15Oct 8, 2026aa_hle
Show 1 more reasoning resultHide 1 reasoning result
- GPQA Diamond73.10Oct 7, 2026GPQA
0 of 10 ranked benchmarks measured
- Terminal-Bench Hard7.58Oct 8, 2026aa_terminalbench_hard
- SciCode28.47Sep 4, 2026aa_scicode
- 65.60Oct 7, 2026
0 of 4 ranked benchmarks measured
- AA-Omniscience · Accuracy16.97Oct 8, 2026omniscienceAccuracy
- AA-Omniscience · Non-hallucination16.32Oct 8, 2026omniscienceNonHallucination
More results
Benchmarks outside the capability baskets. They are not ranked against a leader.
- AA Intelligence11.87Oct 8, 2026aa_intelligence_index
- -52.52Oct 8, 2026
- τ²-Bench Telecom (AA run)45.61Oct 8, 2026aa_tau2
- 14.54Jun 18, 2026
- 23.40Jun 18, 2026
- 73.20Oct 7, 2026
Show 26 more resultsHide 26 results
- 88.90Oct 7, 2026
- 83.70Oct 7, 2026
- 60.50Oct 7, 2026
- 71.70Oct 7, 2026
- 62.80Oct 7, 2026
- 96.10Oct 7, 2026
- 52.30Oct 7, 2026
- 67.40Oct 7, 2026
- 87.80Aug 31, 2026
- 76.30Oct 7, 2026
- 89.20Oct 7, 2026
- 62.60Oct 7, 2026
- 70.20Oct 7, 2026
- 8.30Oct 7, 2026
- 82.10Oct 7, 2026
- 77.20Oct 7, 2026
- 91.90Oct 7, 2026
- MMMU (val) (Pass@1)78.10Oct 7, 2026MMMU (val)
- OCRBench v2 (Chinese)62.10Oct 7, 2026OCRBench-V2 (zh)
- OCRBench v2_en68.40Oct 7, 2026OCRBench-V2 (en)
- 78.40Oct 7, 2026
- 95.70Oct 7, 2026
- ScreenSpot-Pro (No tools)57.10Oct 7, 2026ScreenSpot Pro
- 55.40Oct 7, 2026
- 59.00Oct 7, 2026
- 79.00Oct 7, 2026
Qwen3 VL 32B Reasoning: common questions
Who makes Qwen3 VL 32B Reasoning?
Qwen3 VL 32B Reasoning is made by Alibaba.
When was Qwen3 VL 32B Reasoning released?
Qwen3 VL 32B Reasoning was released on Oct 21, 2025, according to Artificial Analysis.
What is Qwen3 VL 32B Reasoning good at?
Qwen3 VL 32B Reasoning is behind the leaders in multimodal tasks, instruction following, long context, and agentic tasks. Too few results yet to rate reasoning, coding, safety, math, multilingual tasks, or factuality.
How much does Qwen3 VL 32B Reasoning cost?
Qwen3 VL 32B Reasoning costs $0.16 per million input tokens and $0.64 per million output tokens, according to Alibaba's own price page. We track its price at 2 providers. At a mix of three input tokens to one output token, it is cheaper than 73% of the 330 priced models we track.
How many benchmarks has Qwen3 VL 32B Reasoning been tested on?
We track 52 results for Qwen3 VL 32B Reasoning on 50 benchmarks from 2 sources. The latest was recorded on Oct 8, 2026.
About this record
Where Qwen3 VL 32B Reasoning's numbers come from, and every name it appears under.
- Tracked since
- Apr 27, 2026
- Newest source mention
- Sep 10, 2026
Where the results come from
Verification: 52 scores · 0 independently verified · 16 aggregator-attributed · 36 vendor-reported. How these tiers are assigned
From 2 sources on 2 sites. api.llm-stats.com supplies 36 of them. Bars are coloured by trust tier.
- api.llm-stats.com36
- artificialanalysis.ai16