VECTOR WIREAI INTELLIGENCE
PKT
Refresh Models Deals Regulatory Sources

Questions this page answers

Which is better, Qwen2.5 VL 32B Instruct or Qwen3 VL 235B A22B Reasoning?
Across 13 shared benchmarks, Qwen2.5 VL 32B Instruct scores higher on 0 and Qwen3 VL 235B A22B Reasoning on 13. The widest gap is OSWorld-Verified, where Qwen3 VL 235B A22B Reasoning scores 38.1 against 5.9.

Qwen2.5 VL 32B Instruct vs Qwen3 VL 235B A22B Reasoning

Across 13 shared benchmarks, Qwen2.5 VL 32B Instruct scores higher on 0 and Qwen3 VL 235B A22B Reasoning on 13. The widest gap is OSWorld-Verified, where Qwen3 VL 235B A22B Reasoning scores 38.1 against 5.9.

AlibabavsAlibaba13 shared benchmarks013 head-to-head
BenchmarkQwen2.5 VL 32B InstructQwen3 VL 235B A22B Reasoning
arena_vision11531208
CC-OCR77.181.5
GPQA Diamond4677.2
LVBench4963.6
mathvision4074.6
MathVista74.785.8
MMLU-Pro68.883.8
MMMU7078.7
MMMU-Pro49.569.3
MMStar69.578.7
OSWorld-Verified5.938.1
screenspot_pro_no_tools39.461.8
Video-MME77.979

Best tracked score per model per benchmark (default configuration; source-attributed). ↓ marks lower-is-better metrics. Open either model for its full surface, provenance and pricing.