Qwen3 VL 235B A22B Instruct
Qwen3 VL 235B A22B Instruct is capable in multimodal tasks and behind the leaders in agentic tasks and instruction following. Too few results yet to rate reasoning, coding, safety, long context, math, multilingual tasks, or factuality.
Capability profile
Bars show the model's median result as a share of the leading model's, per capability. Select a row to see its results.
Too few results yet to rate Reasoning, Coding, Safety, Long Context, Math, Multilingual or Factuality.
Price
$0.40input$1.60outputper million tokens
From Alibaba's own price page · 5 providers tracked · All prices
Evidence
57results on55benchmarks
- 2 independently verified
- 16 aggregator
- 39 vendor-reported
From 4 sources · latest Oct 8, 2026 · How verification works
API features
Tool callingStructured outputs
As listed by OpenRouter
Qwen3 VL 235B A22B Instruct benchmark results
57 results on 55 benchmarks, grouped by capability. Each bar is the result as a share of the capability leader's; every result links to its source.
23.4% behind the leader4 of 6 ranked benchmarks measured
- 92.00Oct 7, 2026
- MathVista84.90Oct 7, 2026MathVista-Mini
- MMMU-Pro67.57Oct 8, 2026aa_mmmu_pro
- CharXiv (reasoning)62.10Oct 7, 2026CharXiv-R
Show 5 more multimodal resultsHide 5 multimodal results
- 70.70Oct 7, 2026
- 1246.98Aug 25, 2026
- 1215Jun 17, 2026
- 68.10Oct 7, 2026
- 78.40Oct 7, 2026
42.1% behind the leader2 of 7 ranked benchmarks measured
- OSWorld-Verified66.70Oct 7, 2026OSWorld
- 6.73Jun 15, 2026
43.2% behind the leader1 of 3 ranked benchmarks measured
- IFBench42.65Oct 8, 2026aa_ifbench
0 of 6 ranked benchmarks measured
- 0.00Oct 8, 2026
- GPQA Diamond71.21Oct 8, 2026gpqa
- Humanity's Last Exam6.63Oct 8, 2026aa_hle
0 of 10 ranked benchmarks measured
- Terminal-Bench Hard6.82Oct 8, 2026aa_terminalbench_hard
- SciCode35.88Sep 4, 2026aa_scicode
- 54.30Oct 7, 2026
0 of 3 ranked benchmarks measured
- 32.67Oct 8, 2026
0 of 4 ranked benchmarks measured
- AA-Omniscience · Accuracy20.30Oct 8, 2026omniscienceAccuracy
- AA-Omniscience · Non-hallucination8.45Oct 8, 2026omniscienceNonHallucination
More results
Benchmarks outside the capability baskets. They are not ranked against a leader.
- AA Intelligence9.94Oct 8, 2026aa_intelligence_index
- -52.67Oct 8, 2026
- τ²-Bench Telecom (AA run)35.09Oct 8, 2026aa_tau2
- 16.51Jun 18, 2026
- 19.13Jun 18, 2026
- ScreenSpot-Pro (No tools)62.00Oct 7, 2026ScreenSpot Pro
Show 30 more resultsHide 30 results
- 89.70Oct 7, 2026
- 74.70Oct 7, 2026
- 77.40Oct 7, 2026
- 67.70Oct 7, 2026
- 82.20Oct 7, 2026
- 64.80Oct 7, 2026
- 97.10Oct 7, 2026
- 51.30Oct 7, 2026
- 63.20Oct 7, 2026
- HMMT 202557.40Oct 7, 2026HMMT25
- 87.80Aug 31, 2026
- 80.00Oct 7, 2026
- 89.20Oct 7, 2026
- LiveCodeBench (v5)61.40Oct 7, 2026LiveCodeBench v5
- 67.70Oct 7, 2026
- 66.50Oct 7, 2026
- 84.30Oct 7, 2026
- 8.50Oct 7, 2026
- 81.80Oct 7, 2026
- 77.80Oct 7, 2026
- 92.20Oct 7, 2026
- MMMU (val) (Pass@1)78.70Aug 27, 2026MMMUval
- 86.10Oct 7, 2026
- OCRBench v2 (Chinese)61.80Oct 7, 2026OCRBench-V2 (zh)
- OCRBench v2_en67.10Oct 7, 2026OCRBench-V2 (en)
- 79.30Oct 7, 2026
- 95.40Oct 7, 2026
- 51.90Oct 7, 2026
- 60.40Oct 7, 2026
- 74.70Oct 7, 2026
Qwen3 VL 235B A22B Instruct: common questions
Who makes Qwen3 VL 235B A22B Instruct?
Qwen3 VL 235B A22B Instruct is made by Alibaba.
When was Qwen3 VL 235B A22B Instruct released?
Qwen3 VL 235B A22B Instruct was released on Sep 23, 2025, according to Artificial Analysis.
What is Qwen3 VL 235B A22B Instruct good at?
Qwen3 VL 235B A22B Instruct is capable in multimodal tasks and behind the leaders in agentic tasks and instruction following. Too few results yet to rate reasoning, coding, safety, long context, math, multilingual tasks, or factuality.
How much does Qwen3 VL 235B A22B Instruct cost?
Qwen3 VL 235B A22B Instruct costs $0.40 per million input tokens and $1.60 per million output tokens, according to Alibaba's own price page. We track its price at 5 providers. At a mix of three input tokens to one output token, it costs more than 51% of the 330 priced models we track.
How many benchmarks has Qwen3 VL 235B A22B Instruct been tested on?
We track 57 results for Qwen3 VL 235B A22B Instruct on 55 benchmarks from 4 sources, 2 of them independently verified. The latest was recorded on Oct 8, 2026.
Which API features does Qwen3 VL 235B A22B Instruct support?
OpenRouter lists tool calling and structured outputs for Qwen3 VL 235B A22B Instruct.
About this record
Where Qwen3 VL 235B A22B Instruct's numbers come from, and every name it appears under.
- Tracked since
- Apr 27, 2026
- Newest source mention
- May 16, 2026
Where the results come from
Verification: 57 scores · 2 independently verified · 16 aggregator-attributed · 39 vendor-reported. How these tiers are assigned
From 4 sources on 4 sites. api.llm-stats.com supplies 39 of them; the 2 independently verified results come from 2 sites. Bars are coloured by trust tier.
- api.llm-stats.com39
- artificialanalysis.ai16
- datasets-server.huggingface.co1
- lmarena.ai1