MiniMax M2.7
MiniMax M2.7 is capable in long context and coding and behind the leaders in instruction following and reasoning. Too few results yet to rate agentic tasks, safety, math, multimodal tasks, multilingual tasks, or factuality.
Capability profile
Bars show the model's median result as a share of the leading model's, per capability. Select a row to see its results.
Too few results yet to rate Agentic, Safety, Math, Multimodal, Multilingual or Factuality.
Price
$0.30input$1.20outputper million tokens
From Artificial Analysis · 3 providers tracked · All prices
Evidence
99results on79benchmarks
- 5 independently verified
- 20 aggregator
- 8 vendor-reported
- 66 cross-referenced
From 9 sources · latest Oct 8, 2026 · How verification works
API features
Tool callingJSON modeReasoning
As listed by OpenRouter
Research
8 papers reference MiniMax M2.7MiniMax M2.7 benchmark results
99 results on 79 benchmarks, grouped by capability. Each bar is the result as a share of the capability leader's; every result links to its source.
14.0% behind the leader1 of 3 ranked benchmarks measured
- 78.33Oct 8, 2026
23.4% behind the leader8 of 10 ranked benchmarks measured
- LiveCodeBench v677.20Jun 6, 2026LiveCodeBench (v6)
- 79.90Jul 30, 2026
- 76.50Oct 7, 2026
- SciCode50.12Oct 8, 2026aa_scicode
- 56.20Oct 7, 2026
- Terminal-Bench 2.155.43Oct 8, 2026terminalbenchV21
- Terminal-Bench Hard39.39Oct 8, 2026aa_terminalbench_hard
- 1397.69May 22, 2026
Show 6 more coding resultsHide 6 coding results
- 47.00Jul 30, 2026
- 71.80Jun 6, 2026
- 56.20May 18, 2026
- 75.30Jun 6, 2026
- 55.40Jul 30, 2026
- 55.50Jun 6, 2026
29.0% behind the leader2 of 3 ranked benchmarks measured
- IFBench75.71Oct 8, 2026aa_ifbench
- 42.50Jun 10, 2026
Show 1 more instruction following resultHide 1 instruction following result
- 75.70Aug 24, 2026
32.0% behind the leader3 of 6 ranked benchmarks measured
- GPQA Diamond87.37Oct 8, 2026gpqa
- Humanity's Last Exam29.61Oct 8, 2026aa_hle
- 0.57Oct 8, 2026
Show 7 more reasoning resultsHide 7 reasoning results
- 0.60Jul 30, 2026
- GPQA Diamond86.60Aug 24, 2026GPQA (no tools)
- 87.40Jul 30, 2026
- 87.00May 18, 2026
- Humanity's Last Exam28.10Jul 30, 2026HLE text only
- Humanity's Last Exam23.10Jun 6, 2026HLE (no tools)
- 28.00May 18, 2026
0 of 7 ranked benchmarks measured
- 10.62Oct 8, 2026
- 26.46Oct 8, 2026
- 25.67Oct 8, 2026
Show 6 more agentic resultsHide 6 agentic results
- 54.10Jun 6, 2026
- GDPval (win rate)47.60Jun 6, 2026GDPVal
- MCP Atlas48.80May 18, 2026MCP-Atlas (Public Set)
- Terminal-Bench 4.00.00Oct 8, 2026
- τ-Bench V3 · Banking9.90Oct 8, 2026tauBanking
- 14.60Jun 6, 2026
0 of 5 ranked benchmarks measured
- 71.20Jul 30, 2026
- 87.70Jul 30, 2026
- 66.30May 18, 2026
Show 2 more math resultsHide 2 math results
- 89.80May 18, 2026
- HMMT Feb 202672.70May 18, 2026HMMT Feb. 2026
0 of 4 ranked benchmarks measured
- 12.90May 2, 2026
- AA-Omniscience · Accuracy26.80Oct 8, 2026omniscienceAccuracy
- AA-Omniscience · Non-hallucination64.44Oct 8, 2026omniscienceNonHallucination
Show 1 more factuality resultHide 1 factuality result
- 13.50Aug 24, 2026
More results
Benchmarks outside the capability baskets. They are not ranked against a leader.
- vectara_avg_summary_length131.90May 2, 2026Average Summary Length (Words)
- vectara_answer_rate99.40May 2, 2026Answer Rate
- vectara_factual_consistency87.10May 2, 2026Factual Consistency Rate
- AA Intelligence22.76Oct 8, 2026aa_intelligence_index
- 0.77Oct 8, 2026
- τ²-Bench Telecom (AA run)84.80Oct 8, 2026aa_tau2
Show 47 more resultsHide 47 results
- 16.81Sep 9, 2026
- 52.40Jun 24, 2026
- 55.82Jun 24, 2026
- 57.73Jun 24, 2026
- 46.12Jun 24, 2026
- 27.30Jun 24, 2026
- 37.44Jun 24, 2026
- 41.62Jun 24, 2026
- 50.52Jun 24, 2026
- 28.90Jun 6, 2026
- 51.90Jun 6, 2026
- 50.00Oct 7, 2026
- Artificial Analysis Coding Index52.62Sep 9, 2026aa_coding_index
- BrowseComp (context management)76.30Aug 24, 2026BrowseComp (with context management)
- 0.60Jun 6, 2026
- 27.89Oct 7, 2026
- 1159.00Jul 30, 2026
- 83.90Aug 24, 2026
- GPQA (unspecified)86.60Jun 6, 2026GPQA (no tools)
- HLE (with tools)40.30Jul 30, 2026HLE with tools
- 81.00May 18, 2026
- 74.60Jun 10, 2026
- 75.10Jun 6, 2026
- 81.90Jun 6, 2026
- 78.40Jun 10, 2026
- 52.70Oct 7, 2026
- 39.80Oct 7, 2026
- 39.80May 18, 2026
- 77.60Jun 6, 2026
- 52.00Jun 6, 2026
- 38.30Jun 6, 2026
- 99.40Aug 24, 2026
- 56.20Jul 30, 2026
- 75.30Jun 6, 2026
- 66.10Jun 6, 2026
- 84.90Jun 6, 2026
- TauBench V3 - Telecom89.60Aug 24, 2026Telecom
- 57.00Oct 7, 2026
- Terminal-Bench 2.057.00May 18, 2026Terminal-Bench 2.0 (Best self-reported)
- 46.30May 18, 2026
- 46.30Oct 7, 2026
- 47.50Jul 30, 2026
- 50.50Jun 6, 2026
- 51.30Jun 6, 2026
- 82.80Jun 10, 2026
- 67.60May 18, 2026
- τ³-Bench Banking8.90Jul 30, 2026τ³-Banking
MiniMax M2.7: common questions
Who makes MiniMax M2.7?
MiniMax M2.7 is made by MiniMax.
When was MiniMax M2.7 released?
MiniMax M2.7 was released on Mar 18, 2026, according to Artificial Analysis.
What is MiniMax M2.7 good at?
MiniMax M2.7 is capable in long context and coding and behind the leaders in instruction following and reasoning. Too few results yet to rate agentic tasks, safety, math, multimodal tasks, multilingual tasks, or factuality.
How much does MiniMax M2.7 cost?
MiniMax M2.7 costs $0.30 per million input tokens and $1.20 per million output tokens, according to Artificial Analysis. We track its price at 3 providers. At a mix of three input tokens to one output token, it is cheaper than 59% of the 331 priced models we track.
How many benchmarks has MiniMax M2.7 been tested on?
We track 99 results for MiniMax M2.7 on 79 benchmarks from 9 sources, 5 of them independently verified. The latest was recorded on Oct 8, 2026.
Which API features does MiniMax M2.7 support?
OpenRouter lists tool calling, json mode, and reasoning for MiniMax M2.7.
About this record
Where MiniMax M2.7's numbers come from, and every name it appears under.
- Tracked since
- Apr 27, 2026
- Newest source mention
- Sep 3, 2026
Where the results come from
Verification: 99 scores · 5 independently verified · 20 aggregator-attributed · 66 vendor cross-reference · 8 vendor-reported. How these tiers are assigned
From 9 sources on 6 sites. Hugging Face supplies 53 of them; the 5 independently verified results come from 2 sites. Bars are coloured by trust tier.
- huggingface.co53
- artificialanalysis.ai20
- thinkingmachines.ai13
- api.llm-stats.com8
- raw.githubusercontent.com4
- datasets-server.huggingface.co1