Skip to content
VECTOR WIREAI INTELLIGENCE
UTC

MiniMax M2.5

Score basis

MiniMax M2.5 is capable in long context; and behind the leaders in instruction following, coding, agentic tasks, reasoning, and factuality. Too few results yet to rate safety, math, multimodal tasks, or multilingual tasks.

Capability profile

Bars show the model's median result as a share of the leading model's, per capability. Select a row to see its results.

Capable1 capability
  1. Long Context−21.3%1 of 3
Limited5 capabilities
  1. Instruction Following−27.2%1 of 3
  2. Coding−29.4%5 of 10
  3. Agentic−35.1%2 of 7
  4. Reasoning−35.8%4 of 6
  5. Factuality−36.0%3 of 4
Not rated4 capabilities

Too few results yet to rate Safety, Math, Multimodal or Multilingual.

Price

$0.30input$1.20outputper million tokens

From Artificial Analysis · 3 providers tracked · All prices

Cheaper than 59% of 330 priced models · 3:1 input-to-output blend, log scale

Evidence

40results on33benchmarks

  • 11 independently verified
  • 15 aggregator
  • 14 vendor-reported

From 9 sources · latest Oct 8, 2026 · How verification works

API features

Tool callingStructured outputsReasoning

As listed by OpenRouter

MiniMax M2.5 benchmark results

40 results on 33 benchmarks, grouped by capability. Each bar is the result as a share of the capability leader's; every result links to its source.

Long Context

CapableFull long context ranking

21.3% behind the leader1 of 3 ranked benchmarks measured

Instruction Following

LimitedFull instruction following ranking

27.2% behind the leader1 of 3 ranked benchmarks measured

Show 1 more instruction following resultHide 1 instruction following result

Coding

LimitedFull coding ranking

29.4% behind the leader5 of 10 ranked benchmarks measured

Show 5 more coding resultsHide 5 coding results

Agentic

LimitedFull agentic ranking

35.1% behind the leader2 of 7 ranked benchmarks measured

Reasoning

LimitedFull reasoning ranking

35.8% behind the leader4 of 6 ranked benchmarks measured

Show 2 more reasoning resultsHide 2 reasoning results

Factuality

LimitedFull factuality ranking

36.0% behind the leader3 of 4 ranked benchmarks measured

More results

Benchmarks outside the capability baskets. They are not ranked against a leader.

Show 10 more resultsHide 10 results

MiniMax M2.5: common questions

Who makes MiniMax M2.5?

MiniMax M2.5 is made by MiniMax.

When was MiniMax M2.5 released?

MiniMax M2.5 was released on Feb 12, 2026, according to Artificial Analysis.

What is MiniMax M2.5 good at?

MiniMax M2.5 is capable in long context; and behind the leaders in instruction following, coding, agentic tasks, reasoning, and factuality. Too few results yet to rate safety, math, multimodal tasks, or multilingual tasks.

How much does MiniMax M2.5 cost?

MiniMax M2.5 costs $0.30 per million input tokens and $1.20 per million output tokens, according to Artificial Analysis. We track its price at 3 providers. At a mix of three input tokens to one output token, it is cheaper than 59% of the 330 priced models we track.

How many benchmarks has MiniMax M2.5 been tested on?

We track 40 results for MiniMax M2.5 on 33 benchmarks from 9 sources, 11 of them independently verified. The latest was recorded on Oct 8, 2026.

Which API features does MiniMax M2.5 support?

OpenRouter lists tool calling, structured outputs, and reasoning for MiniMax M2.5.

About this record

Where MiniMax M2.5's numbers come from, and every name it appears under.

Tracked since
Apr 27, 2026
Newest source mention
Aug 10, 2026

Where the results come from

Verification: 40 scores · 11 independently verified · 15 aggregator-attributed · 14 vendor-reported. How these tiers are assigned

From 9 sources on 8 sites. Artificial Analysis supplies 15 of them; the 11 independently verified results come from 5 sites. Bars are coloured by trust tier.

  • artificialanalysis.ai15
  • huggingface.co10
  • api.llm-stats.com4
  • raw.githubusercontent.com4
  • swebench.com3
  • arcprize.org2
  • datasets-server.huggingface.co1
  • labs.scale.com1

Also known as

minimax-m2-5minimax m2.5 (high reasoning)mini-swe-agent + minimax m2.5 (high reasoning)minimax 2.5minimax-m2p5minimax-m2.5:freeminimax m2.5 (high)