VECTOR WIREAI INTELLIGENCE
UTC
Refresh Models Deals Regulatory Sources

InternVL-3-8B vs Phi 4 Multimodal Instruct

8 SHARED BENCHMARKS

Across 8 shared benchmarks, InternVL-3-8B scores higher on 6 and Phi 4 Multimodal Instruct on 2. The widest gap is TextVQA-val, where InternVL-3-8B scores 82.2 against 39.9.

OPENGVLABVSMICROSOFT8 SHARED62 HEAD-TO-HEAD
AI2D85.182.3
BLINK55.961.3
ChartQAPro37.30.1
MathVista69.562.4
MMStar68.461.2
POPE90.685.6
TextVQA-val82.239.9

Best tracked score per model per benchmark (default configuration; source-attributed). ↓ marks lower-is-better metrics. Open either model for its full surface, provenance and pricing.