Across 12 shared benchmarks, Qwen2.5 VL 72B scores higher on 8 and Qwen2 VL 72B Instruct on 4. The widest gap is OCRBench, where Qwen2 VL 72B Instruct scores 877 against 88.5.
| Benchmark | Qwen2.5 VL 72B | Qwen2 VL 72B Instruct |
|---|---|---|
| CC-OCR | 79.8 | 68.7 |
| chartqa | 89.5 | 88.3 |
| docvqa | 96.4 | 96.5 |
| EgoSchema | 76.2 | 77.9 |
| mathvision | 38.1 | 25.9 |
| MathVista | 74.8 | 70.5 |
| MMMU | 70.2 | 64.5 |
| MMMU-Pro | 51.1 | 46.2 |
| MMStar | 70.8 | 68.3 |
| MV-Bench | 70.4 | 73.6 |
| OCRBench | 88.5 | 877 |
| Video-MME | 79.1 | 77.8 |
Best tracked score per model per benchmark (default configuration; source-attributed). ↓ marks lower-is-better metrics. Open either model for its full surface, provenance and pricing.