Skip to content
VECTOR WIREAI INTELLIGENCE
UTC

Mistral Small 4

Score basis

Mistral Small 4 is behind the leaders in factuality, multimodal tasks, coding, instruction following, and reasoning. Too few results yet to rate agentic tasks, safety, long context, math, or multilingual tasks.

Capability profile

Bars show the model's median result as a share of the leading model's, per capability. Select a row to see its results.

Limited5 capabilities
  1. Factuality−31.3%2 of 4
  2. Multimodal−37.6%1 of 6
  3. Coding−39.0%3 of 10
  4. Instruction Following−40.0%1 of 3
  5. Reasoning−44.4%3 of 6
Not rated5 capabilities

Too few results yet to rate Agentic, Safety, Long Context, Math or Multilingual.

Price

$0.15input$0.60outputper million tokens

From Artificial Analysis · 2 providers tracked · All prices

Cheaper than 77% of 331 priced models · 3:1 input-to-output blend, log scale

Evidence

43results on23benchmarks

  • 1 independently verified
  • 34 aggregator
  • 8 vendor-reported

From 3 sources · latest Oct 8, 2026 · How verification works

Mistral Small 4 benchmark results

43 results on 23 benchmarks, grouped by capability. Each bar is the result as a share of the capability leader's; every result links to its source.

Factuality

LimitedFull factuality ranking

31.3% behind the leader2 of 4 ranked benchmarks measured

Show 2 more factuality resultsHide 2 factuality results

Multimodal

LimitedFull multimodal ranking

37.6% behind the leader1 of 6 ranked benchmarks measured

Show 2 more multimodal resultsHide 2 multimodal results

Coding

LimitedFull coding ranking

39.0% behind the leader3 of 10 ranked benchmarks measured

Show 2 more coding resultsHide 2 coding results

Instruction Following

LimitedFull instruction following ranking

40.0% behind the leader1 of 3 ranked benchmarks measured

Show 2 more instruction following resultsHide 2 instruction following results

Reasoning

LimitedFull reasoning ranking

44.4% behind the leader3 of 6 ranked benchmarks measured

Show 3 more reasoning resultsHide 3 reasoning results

Agentic

Not enough dataFull agentic ranking

0 of 7 ranked benchmarks measured

Show 1 more agentic resultHide 1 agentic result

Long Context

Not enough dataFull long context ranking

0 of 3 ranked benchmarks measured

More results

Benchmarks outside the capability baskets. They are not ranked against a leader.

Show 9 more resultsHide 9 results

Mistral Small 4: common questions

Who makes Mistral Small 4?

Mistral Small 4 is made by Mistral.

When was Mistral Small 4 released?

Mistral Small 4 was released on Mar 16, 2026, according to Mistral's own announcement.

What is Mistral Small 4 good at?

Mistral Small 4 is behind the leaders in factuality, multimodal tasks, coding, instruction following, and reasoning. Too few results yet to rate agentic tasks, safety, long context, math, or multilingual tasks.

How much does Mistral Small 4 cost?

Mistral Small 4 costs $0.15 per million input tokens and $0.60 per million output tokens, according to Artificial Analysis. We track its price at 2 providers. At a mix of three input tokens to one output token, it is cheaper than 77% of the 331 priced models we track.

How many benchmarks has Mistral Small 4 been tested on?

We track 43 results for Mistral Small 4 on 23 benchmarks from 3 sources, 1 of them independently verified. The latest was recorded on Oct 8, 2026.

About this record

Where Mistral Small 4's numbers come from, and every name it appears under.

Tracked since
Apr 25, 2026
Newest source mention
Jul 20, 2026

Where the results come from

Verification: 43 scores · 1 independently verified · 34 aggregator-attributed · 8 vendor-reported. How these tiers are assigned

From 3 sources on 2 sites. Artificial Analysis supplies 35 of them; the 1 independently verified result comes from 1 site. Bars are coloured by trust tier.

  • artificialanalysis.ai35
  • api.llm-stats.com8

Also known as

mistral small 4 (non-reasoning)Mistral Small 4 (Reasoning)mistral-small-4-non-reasoning