VECTOR WIREAI INTELLIGENCE
UTC
Refresh Models Deals Regulatory Sources

InternVL-3-78B vs Phi 4 Multimodal Instruct

8 SHARED BENCHMARKS

Across 8 shared benchmarks, InternVL-3-78B scores higher on 6 and Phi 4 Multimodal Instruct on 2. The widest gap is TextVQA-val, where InternVL-3-78B scores 83.5 against 39.9.

OPENGVLABVSMICROSOFT8 SHARED62 HEAD-TO-HEAD
AI2D83.582.3
BLINK51.961.3
ChartQAPro44.40.1
DocVQA-val83.892.8
MathVista70.162.4
MMStar66.161.2
POPE88.985.6
TextVQA-val83.539.9

Best tracked score per model per benchmark (default configuration; source-attributed). ↓ marks lower-is-better metrics. Open either model for its full surface, provenance and pricing.