Skip to content
VECTOR WIREAI INTELLIGENCE
UTC

Grok 3 Mini Reasoning

Score basis

Grok 3 Mini Reasoning is behind the leaders in long context, instruction following, and reasoning. Too few results yet to rate coding, agentic tasks, safety, math, multimodal tasks, multilingual tasks, or factuality.

Capability profile

Bars show the model's median result as a share of the leading model's, per capability. Select a row to see its results.

Limited3 capabilities
  1. Long Context−34.0%1 of 3
  2. Instruction Following−40.7%1 of 3
  3. Reasoning−45.3%4 of 6
Not rated7 capabilities

Too few results yet to rate Coding, Agentic, Safety, Math, Multimodal, Multilingual or Factuality.

Price

$0.30input$0.50outputper million tokens

From Artificial Analysis · All prices

Cheaper than 67% of 331 priced models · 3:1 input-to-output blend, log scale

Evidence

30results on25benchmarks

  • 7 independently verified
  • 15 aggregator
  • 8 vendor-reported

From 7 sources · latest Oct 8, 2026 · How verification works

Grok 3 Mini Reasoning benchmark results

30 results on 25 benchmarks, grouped by capability. Each bar is the result as a share of the capability leader's; every result links to its source.

Long Context

LimitedFull long context ranking

34.0% behind the leader1 of 3 ranked benchmarks measured

Instruction Following

LimitedFull instruction following ranking

40.7% behind the leader1 of 3 ranked benchmarks measured

Reasoning

LimitedFull reasoning ranking

45.3% behind the leader4 of 6 ranked benchmarks measured

Show 3 more reasoning resultsHide 3 reasoning results

Coding

Not enough dataFull coding ranking

0 of 10 ranked benchmarks measured

Agentic

Not enough dataFull agentic ranking

0 of 7 ranked benchmarks measured

Factuality

Not enough dataFull factuality ranking

0 of 4 ranked benchmarks measured

More results

Benchmarks outside the capability baskets. They are not ranked against a leader.

Show 9 more resultsHide 9 results

Grok 3 Mini Reasoning: common questions

Who makes Grok 3 Mini Reasoning?

Grok 3 Mini Reasoning is made by SpaceXAI.

When was Grok 3 Mini Reasoning released?

Grok 3 Mini Reasoning was released on Feb 19, 2025, according to Artificial Analysis.

What is Grok 3 Mini Reasoning good at?

Grok 3 Mini Reasoning is behind the leaders in long context, instruction following, and reasoning. Too few results yet to rate coding, agentic tasks, safety, math, multimodal tasks, multilingual tasks, or factuality.

How much does Grok 3 Mini Reasoning cost?

Grok 3 Mini Reasoning costs $0.30 per million input tokens and $0.50 per million output tokens, according to Artificial Analysis. At a mix of three input tokens to one output token, it is cheaper than 67% of the 331 priced models we track.

How many benchmarks has Grok 3 Mini Reasoning been tested on?

We track 30 results for Grok 3 Mini Reasoning on 25 benchmarks from 7 sources, 7 of them independently verified. The latest was recorded on Oct 8, 2026.

About this record

Where Grok 3 Mini Reasoning's numbers come from, and every name it appears under.

Tracked since
Apr 27, 2026
Newest source mention
Jul 13, 2026

Where the results come from

Verification: 30 scores · 7 independently verified · 15 aggregator-attributed · 8 vendor-reported. How these tiers are assigned

From 7 sources on 6 sites. Artificial Analysis supplies 15 of them; the 7 independently verified results come from 3 sites. Bars are coloured by trust tier.

  • artificialanalysis.ai15
  • x.ai5
  • livecodebench.github.io4
  • api.llm-stats.com3
  • arcprize.org2
  • epoch.ai1

Also known as

How our sources name Grok 3 Mini Reasoning at each reasoning setting.

SettingShort formAPI id
lowGrok 3 Mini (Low)—
highgrok 3 mini reasoning (high)grok-3-mini (high)
Also listed asgrok-3 mini