Molmo2-8B
Molmo2-8B is behind the leaders in multimodal tasks. Too few results yet to rate reasoning, coding, agentic tasks, safety, long context, math, multilingual tasks, instruction following, or factuality.
Capability profile
Bars show the model's median result as a share of the leading model's, per capability. Select a row to see its results.
Too few results yet to rate Reasoning, Coding, Agentic, Safety, Long Context, Math, Multilingual, Instruction Following or Factuality.
Price
No current price is tracked for this model. See the rate card
Evidence
14results on14benchmarks
- 1 independently verified
- 13 aggregator
From 2 sources · latest Oct 7, 2026 · How verification works
Molmo2-8B benchmark results
14 results on 14 benchmarks, grouped by capability. Each bar is the result as a share of the capability leader's; every result links to its source.
42.9% behind the leader1 of 6 ranked benchmarks measured
- MMMU-Pro37.46Oct 7, 2026aa_mmmu_pro
Show 1 more multimodal resultHide 1 multimodal result
- 44.70May 22, 2026
0 of 6 ranked benchmarks measured
- 0.00Oct 7, 2026
- GPQA Diamond42.53Oct 7, 2026gpqa
- Humanity's Last Exam4.26Oct 7, 2026aa_hle
0 of 10 ranked benchmarks measured
- Terminal-Bench Hard0.00Oct 7, 2026aa_terminalbench_hard
- SciCode13.31Sep 4, 2026aa_scicode
0 of 3 ranked benchmarks measured
- 0.00Oct 7, 2026
0 of 3 ranked benchmarks measured
- IFBench26.94Oct 7, 2026aa_ifbench
0 of 4 ranked benchmarks measured
- AA-Omniscience · Accuracy11.45Oct 7, 2026omniscienceAccuracy
- AA-Omniscience · Non-hallucination8.90Oct 7, 2026omniscienceNonHallucination
More results
Benchmarks outside the capability baskets. They are not ranked against a leader.
- τ²-Bench Telecom (AA run)0.00Oct 7, 2026aa_tau2
- -69.22Oct 7, 2026
- AA Intelligence5.00Oct 7, 2026aa_intelligence_index
Molmo2-8B: common questions
Who makes Molmo2-8B?
Molmo2-8B is made by AllenAI.
When was Molmo2-8B released?
Molmo2-8B was released on Dec 11, 2025, according to Artificial Analysis.
What is Molmo2-8B good at?
Molmo2-8B is behind the leaders in multimodal tasks. Too few results yet to rate reasoning, coding, agentic tasks, safety, long context, math, multilingual tasks, instruction following, or factuality.
How many benchmarks has Molmo2-8B been tested on?
We track 14 results for Molmo2-8B on 14 benchmarks from 2 sources, 1 of them independently verified. The latest was recorded on Oct 7, 2026.
About this record
Where Molmo2-8B's numbers come from, and every name it appears under.
- Tracked since
- May 22, 2026
- Newest source mention
- May 22, 2026
Where the results come from
Verification: 14 scores · 1 independently verified · 13 aggregator-attributed. How these tiers are assigned
From 2 sources on 2 sites. Artificial Analysis supplies 13 of them; the 1 independently verified result comes from 1 site. Bars are coloured by trust tier.
- artificialanalysis.ai13
- 99franklin.github.io1