Kimi VL A3B Thinking
Kimi VL A3B Thinking is behind the leaders in multimodal tasks. Too few results yet to rate reasoning, coding, agentic tasks, safety, long context, math, multilingual tasks, instruction following, or factuality.
Capability profile
Bars show the model's median result as a share of the leading model's, per capability. Select a row to see its results.
Too few results yet to rate Reasoning, Coding, Agentic, Safety, Long Context, Math, Multilingual, Instruction Following or Factuality.
Price
No current price is tracked for this model. See the rate card
Evidence
36results on22benchmarks
- 36 vendor-reported
From 2 sources · latest Jun 12, 2026 · How verification works
Kimi VL A3B Thinking benchmark results
36 results on 22 benchmarks, grouped by capability. Each bar is the result as a share of the capability leader's; every result links to its source.
46.4% behind the leader2 of 6 ranked benchmarks measured
0 of 6 ranked benchmarks measured
- 42.30Jun 5, 2026
More results
Benchmarks outside the capability baskets. They are not ranked against a leader.
- 64.20Jun 12, 2026
- 864.00Jun 12, 2026
- 64.00Jun 12, 2026
- 76.00Jun 12, 2026
- 61.70Jun 5, 2026
- 36.80Jun 5, 2026
Show 22 more resultsHide 22 results
- 91.80Jun 5, 2026
- MATH-Vision56.90Jun 5, 2026MATH-Vision (Pass@1)
- MATH-Vision36.80Jun 5, 2026MATH-Vision (Pass@1)
- 84.40Jun 5, 2026
- 42.10Jun 5, 2026
- 32.50Jun 5, 2026
- 82.00Jun 5, 2026
- 68.50Jun 5, 2026
- 70.40Jun 5, 2026
- 78.10Jun 5, 2026
- 69.50Jun 5, 2026
- MMVU57.50Jun 5, 2026MMVU (Pass@1)
- MMVU53.00Jun 5, 2026MMVU (Pass@1)
- 869.00Jun 5, 2026
- 52.50Jun 5, 2026
- 70.00Jun 5, 2026
- 52.80Jun 5, 2026
- ScreenSpot-V291.40Jun 5, 2026ScreenSpot-V2 (Acc)
- VideoMME (w sub.)71.90Jun 5, 2026Video-MME (w/ sub.)
- VideoMME (w sub.)66.00Jun 5, 2026Video-MME (w/ sub.)
- VideoMMMU65.20Jun 5, 2026VideoMMMU (Pass@1)
- VideoMMMU55.50Jun 5, 2026VideoMMMU (Pass@1)
Kimi VL A3B Thinking: common questions
Who makes Kimi VL A3B Thinking?
Kimi VL A3B Thinking is made by Moonshot.
When was Kimi VL A3B Thinking released?
Kimi VL A3B Thinking's weights were first published on Hugging Face on Apr 9, 2025.
What is Kimi VL A3B Thinking good at?
Kimi VL A3B Thinking is behind the leaders in multimodal tasks. Too few results yet to rate reasoning, coding, agentic tasks, safety, long context, math, multilingual tasks, instruction following, or factuality.
How many benchmarks has Kimi VL A3B Thinking been tested on?
We track 36 results for Kimi VL A3B Thinking on 22 benchmarks from 2 sources. The latest was recorded on Jun 12, 2026.
About this record
Where Kimi VL A3B Thinking's numbers come from, and every name it appears under.
- Tracked since
- Jun 5, 2026
- Newest source mention
- Jun 5, 2026
Where the results come from
Verification: 36 scores · 0 independently verified · 36 vendor-reported. How these tiers are assigned
From 2 sources on 1 site. Hugging Face supplies all of them. Bars are coloured by trust tier.
- huggingface.co36