Phi 4 Multimodal Instruct
Phi 4 Multimodal Instruct is behind the leaders in multimodal tasks. Too few results yet to rate reasoning, coding, agentic tasks, safety, long context, math, multilingual tasks, instruction following, or factuality.
Capability profile
Bars show the model's median result as a share of the leading model's, per capability. Select a row to see its results.
Too few results yet to rate Reasoning, Coding, Agentic, Safety, Long Context, Math, Multilingual, Instruction Following or Factuality.
Price
No current price is tracked for this model. See the rate card
Evidence
44results on39benchmarks
- 1 independently verified
- 5 aggregator
- 38 vendor-reported
From 4 sources · latest Oct 7, 2026 · How verification works
Phi 4 Multimodal Instruct benchmark results
44 results on 39 benchmarks, grouped by capability. Each bar is the result as a share of the capability leader's; every result links to its source.
59.9% behind the leader5 of 6 ranked benchmarks measured
- 84.40Oct 6, 2026
- 62.40Oct 6, 2026
- 55.10Oct 6, 2026
- 55.00Oct 6, 2026
- MMMU-Pro14.51Oct 7, 2026aa_mmmu_pro
0 of 6 ranked benchmarks measured
- GPQA Diamond31.52Oct 7, 2026gpqa
- Humanity's Last Exam5.00Oct 7, 2026aa_hle
0 of 10 ranked benchmarks measured
- SciCode11.00Sep 4, 2026aa_scicode
More results
Benchmarks outside the capability baskets. They are not ranked against a leader.
- AA Intelligence5.81Oct 7, 2026aa_intelligence_index
- 75.60Oct 6, 2026
- 85.60Oct 6, 2026
- 86.70Oct 6, 2026
- 72.70Oct 6, 2026
- 93.20Oct 6, 2026
Show 23 more resultsHide 23 results
- 82.30Oct 6, 2026
- 81.83Jul 26, 2026
- 81.40Oct 6, 2026
- 83.76Jul 26, 2026
- 0.13Jul 26, 2026
- 57.09Jul 26, 2026
- 92.79Jul 26, 2026
- InfoVQA (val)71.84Jul 26, 2026InfoVQA-val
- 41.14Jul 26, 2026
- 25.31Jul 26, 2026
- 44.18Jul 26, 2026
- 75.17Jul 26, 2026
- 65.81Jul 26, 2026
- 1409.66Jul 26, 2026
- 32.45Jul 26, 2026
- 46.84Jul 26, 2026
- 44.90Jul 26, 2026
- 54.10Jul 26, 2026
- 3.14Jul 26, 2026
- 55.33Jul 26, 2026
- 39.93Jul 26, 2026
- 2.03Jul 26, 2026
- 24.09Jul 26, 2026
Phi 4 Multimodal Instruct: common questions
Who makes Phi 4 Multimodal Instruct?
Phi 4 Multimodal Instruct is made by Microsoft.
When was Phi 4 Multimodal Instruct released?
Phi 4 Multimodal Instruct was released on Feb 26, 2025, according to Artificial Analysis.
What is Phi 4 Multimodal Instruct good at?
Phi 4 Multimodal Instruct is behind the leaders in multimodal tasks. Too few results yet to rate reasoning, coding, agentic tasks, safety, long context, math, multilingual tasks, instruction following, or factuality.
How many benchmarks has Phi 4 Multimodal Instruct been tested on?
We track 44 results for Phi 4 Multimodal Instruct on 39 benchmarks from 4 sources, 1 of them independently verified. The latest was recorded on Oct 7, 2026.
About this record
Where Phi 4 Multimodal Instruct's numbers come from, and every name it appears under.
- Tracked since
- May 2, 2026
- Newest source mention
- Sep 10, 2026
Where the results come from
Verification: 44 scores · 1 independently verified · 5 aggregator-attributed · 38 vendor-reported. How these tiers are assigned
From 4 sources on 4 sites. Hugging Face supplies 25 of them; the 1 independently verified result comes from 1 site. Bars are coloured by trust tier.
- huggingface.co25
- api.llm-stats.com13
- artificialanalysis.ai5
- 99franklin.github.io1