VECTOR WIREAI INTELLIGENCE
UTC
Refresh Models Deals Regulatory Sources

InternVL-3-78B vs Qwen3 VL 4B (Reasoning)

11 SHARED BENCHMARKS

Across 11 shared benchmarks, InternVL-3-78B scores higher on 3 and Qwen3 VL 4B (Reasoning) on 8. The widest gap is mathvision, where Qwen3 VL 4B (Reasoning) scores 60 against 34.8.

OPENGVLABVSALIBABA11 SHARED38 HEAD-TO-HEAD
AI2D83.584.9
BLINK51.965.1
ChartQAPro44.436.2
DocVQA-val83.894.7
MathVista70.179.5
MMStar66.173.2
RealWorldQA74.373.2
TextVQA-val83.580.5

Best tracked score per model per benchmark (default configuration; source-attributed). ↓ marks lower-is-better metrics. Open either model for its full surface, provenance and pricing.

Compare Qwen3 VL 4B (Reasoning) withEVERY TRACKED PAIRING

Anthropic12 PAIRINGS
OpenAI11 PAIRINGS
Google11 PAIRINGS
Alibaba10 PAIRINGS
DeepSeek4 PAIRINGS
Moonshot4 PAIRINGS
Meta2 PAIRINGS
Z.ai2 PAIRINGS
MiniMax1 PAIRING
NVIDIA2 PAIRINGS