Skip to content
VECTOR WIREAI INTELLIGENCE
UTC

Phi 4 Multimodal Instruct vs Qwen3 VL 2B

Microsoft · released

Alibaba

Scores updated · 6 tests both models report · How we compare

Every test, side by side

All 6 tests both models report. The winning score is in its model's colour; marks a score checked independently.

Other results6 tests
  • VSI-BenchQwen3 VL 2B by 23.924.148+23.9
  • CV-BenchQwen3 VL 2B by 22.957.180+22.9
  • SATPhi 4 Multimodal Instruct by 1055.345.3+10
  • TextVQAQwen3 VL 2B by 4.375.679.9+4.3
  • ChartQAPhi 4 Multimodal Instruct by 3.181.478.3+3.1
  • DocVQAtie93.292.7tie

Questions people ask

How do you compare the two?

We use the 6 benchmark tests both models have published scores on. None of them is in the eight capability areas we count, so this page lists them without a verdict. Each score is the one shown on the model's own page.