Skip to content
VECTOR WIREAI INTELLIGENCE
UTC

GPT-4o mini

Score basis

GPT-4o mini is behind the leaders in multimodal tasks. Too few results yet to rate reasoning, coding, agentic tasks, safety, long context, math, multilingual tasks, instruction following, or factuality.

Capability profile

Bars show the model's median result as a share of the leading model's, per capability. Select a row to see its results.

Limited1 capability
  1. Multimodal−52.3%3 of 6
Not rated9 capabilities

Too few results yet to rate Reasoning, Coding, Agentic, Safety, Long Context, Math, Multilingual, Instruction Following or Factuality.

Price

$0.15input$0.60outputper million tokens

From Artificial Analysis · 2 providers tracked · All prices

Cheaper than 77% of 330 priced models · 3:1 input-to-output blend, log scale

Evidence

93results on74benchmarks

  • 32 independently verified
  • 11 aggregator
  • 7 vendor-reported
  • 43 cross-referenced

From 28 sources · latest Oct 8, 2026 · How verification works

API features

Tool callingStructured outputsWeb search

As listed by OpenRouter

GPT-4o mini benchmark results

93 results on 74 benchmarks, grouped by capability. Each bar is the result as a share of the capability leader's; every result links to its source.

Multimodal

LimitedFull multimodal ranking

52.3% behind the leader3 of 6 ranked benchmarks measured

Show 4 more multimodal resultsHide 4 multimodal results

Reasoning

Not enough dataFull reasoning ranking

0 of 6 ranked benchmarks measured

Show 4 more reasoning resultsHide 4 reasoning results

Coding

Not enough dataFull coding ranking

0 of 10 ranked benchmarks measured

Show 2 more coding resultsHide 2 coding results

Agentic

Not enough dataFull agentic ranking

0 of 7 ranked benchmarks measured

Math

Not enough dataFull math ranking

0 of 5 ranked benchmarks measured

Instruction Following

Not enough dataFull instruction following ranking

0 of 3 ranked benchmarks measured

Factuality

Not enough dataFull factuality ranking

0 of 4 ranked benchmarks measured

More results

Benchmarks outside the capability baskets. They are not ranked against a leader.

Show 61 more resultsHide 61 results

GPT-4o mini: common questions

Who makes GPT-4o mini?

GPT-4o mini is made by OpenAI.

When was GPT-4o mini released?

GPT-4o mini was released on Jul 18, 2024, according to Artificial Analysis.

What is GPT-4o mini good at?

GPT-4o mini is behind the leaders in multimodal tasks. Too few results yet to rate reasoning, coding, agentic tasks, safety, long context, math, multilingual tasks, instruction following, or factuality.

How much does GPT-4o mini cost?

GPT-4o mini costs $0.15 per million input tokens and $0.60 per million output tokens, according to Artificial Analysis. We track its price at 2 providers. At a mix of three input tokens to one output token, it is cheaper than 77% of the 330 priced models we track.

How many benchmarks has GPT-4o mini been tested on?

We track 93 results for GPT-4o mini on 74 benchmarks from 28 sources, 32 of them independently verified. The latest was recorded on Oct 8, 2026.

Which API features does GPT-4o mini support?

OpenRouter lists tool calling, structured outputs, and web search for GPT-4o mini.

About this record

Where GPT-4o mini's numbers come from, and every name it appears under.

Tracked since
May 2, 2026
Newest source mention
Aug 28, 2026

Where the results come from

Verification: 93 scores · 32 independently verified · 11 aggregator-attributed · 43 vendor cross-reference · 7 vendor-reported. How these tiers are assigned

From 28 sources on 9 sites. Hugging Face supplies 53 of them; the 32 independently verified results come from 8 sites. Bars are coloured by trust tier.

  • huggingface.co53
  • artificialanalysis.ai12
  • storage.googleapis.com12
  • api.llm-stats.com7
  • livecodebench.github.io4
  • epoch.ai2
  • aider.chat1
  • arcprize.org1
  • simple-bench.com1

Also known as

gpt-4o mini (2024-07-18)GPT-4o mini (Non-Reasoning)GPT-4o mini (Reasoning)GPT-4o Mini:batchgpt-4o-mini-2024-07-18gpt-4o-mini 2024-7-18