Skip to content
VECTOR WIREAI INTELLIGENCE
UTC

MiMo V2.5

Basis

MiMo V2.5 is capable in long context and multimodal tasks; and behind the leaders in coding, reasoning, instruction following, and agentic tasks. Too few results yet to rate safety, math, multilingual tasks, or factuality.

Price per million tokens

$0.14input$0.28output

Price from Artificial Analysis · 4 providers tracked · All prices

Cheaper than 82% of 320 priced models · 3:1 input-to-output blend, log scale
Evidence

51results on39benchmarks

  • 3 independently verified
  • 17 aggregator
  • 8 vendor-reported
  • 23 cross-referenced

From 7 sources · latest Sep 27, 2026 · How verification works

API features

Tool callingStructured outputsReasoning

As listed by OpenRouter

Capability profile

Bars show the model's median result as a share of the leading model's, per capability.

Capable2 capabilities
  1. Long Context−20.8%1 of 3
  2. Multimodal−22.8%3 of 6
Limited4 capabilities
  1. Coding−27.5%6 of 10
  2. Reasoning−27.6%3 of 6
  3. Instruction Following−30.7%1 of 3
  4. Agentic−46.0%3 of 7

Too few results to rate: SafetyMathMultilingualFactuality

MiMo V2.5 benchmark results

51 results on 39 benchmarks, grouped by capability. Every score links to its source; a bar is the result as a share of the capability leader's.

Long Context

Capable1 of 3 ranked benchmarks measuredFull long context ranking
  • AA-LCR
    aa_lcr
    73.0082%
    Reasoning
    artificialanalysis.ai
    AggregatorSep 27, 2026view ↗

Multimodal

Capable3 of 6 ranked benchmarks measuredFull multimodal ranking
4 more multimodal results

Coding

Limited6 of 10 ranked benchmarks measuredFull coding ranking
2 more coding results

Reasoning

Limited3 of 6 ranked benchmarks measuredFull reasoning ranking
4 more reasoning results

Instruction Following

Limited1 of 3 ranked benchmarks measuredFull instruction following ranking
  • IFBench
    aa_ifbench
    67.1481%
    Reasoning
    artificialanalysis.ai
    AggregatorSep 27, 2026view ↗
1 more instruction following result

Agentic

Limited3 of 7 ranked benchmarks measuredFull agentic ranking

Math

Not enough data0 of 5 ranked benchmarks measuredFull math ranking

Factuality

Not enough data0 of 4 ranked benchmarks measuredFull factuality ranking

More results

12 more results

About this record

Verification: 51 scores · 3 independently verified · 17 aggregator-attributed · 23 vendor cross-reference · 8 vendor-reported. How these tiers are assigned

Tracked since
Sep 23, 2026
Newest source mention
Sep 23, 2026