VECTOR WIREAI INTELLIGENCE
UTC
Refresh Models Deals Regulatory Sources

InternVL-3-8B vs Qwen2.5 VL 3B

8 SHARED BENCHMARKS

Across 8 shared benchmarks, InternVL-3-8B scores higher on 6 and Qwen2.5 VL 3B on 2. The widest gap is OCRVQA_TEST, where Qwen2.5 VL 3B scores 69.2 against 39.

OPENGVLABVSALIBABA8 SHARED62 HEAD-TO-HEAD
AI2D85.177.1
BLINK55.949
MMStar68.456.1
OCRVQA_TEST3969.2
POPE90.686.2
RealWorldQA71.165.2
TextVQA-val82.279.1

Best tracked score per model per benchmark (default configuration; source-attributed). ↓ marks lower-is-better metrics. Open either model for its full surface, provenance and pricing.