Skip to content
VECTOR WIREAI INTELLIGENCE
UTC

GPT-5.1 Codex mini

Score basis

GPT-5.1 Codex mini is behind the leaders in factuality, long context, instruction following, multimodal tasks, and agentic tasks. Too few results yet to rate reasoning, coding, safety, math, or multilingual tasks.

Capability profile

Bars show the model's median result as a share of the leading model's, per capability. Select a row to see its results.

Limited5 capabilities
  1. Factuality−27.6%2 of 4
  2. Long Context−27.7%1 of 3
  3. Instruction Following−30.2%1 of 3
  4. Multimodal−32.8%1 of 6
  5. Agentic−37.7%1 of 7
Not rated5 capabilities

Too few results yet to rate Reasoning, Coding, Safety, Math or Multilingual.

Price

$0.25input$2.00outputper million tokens

From Artificial Analysis · 2 providers tracked · All prices

Cheaper than 50% of 330 priced models · 3:1 input-to-output blend, log scale

Evidence

18results on18benchmarks

  • 1 independently verified
  • 16 aggregator
  • 1 vendor-reported

From 3 sources · latest Oct 8, 2026 · How verification works

API features

Tool callingStructured outputsReasoning

As listed by OpenRouter

GPT-5.1 Codex mini benchmark results

18 results on 18 benchmarks, grouped by capability. Each bar is the result as a share of the capability leader's; every result links to its source.

Factuality

LimitedFull factuality ranking

27.6% behind the leader2 of 4 ranked benchmarks measured

Long Context

LimitedFull long context ranking

27.7% behind the leader1 of 3 ranked benchmarks measured

Instruction Following

LimitedFull instruction following ranking

30.2% behind the leader1 of 3 ranked benchmarks measured

Multimodal

LimitedFull multimodal ranking

32.8% behind the leader1 of 6 ranked benchmarks measured

Agentic

LimitedFull agentic ranking

37.7% behind the leader1 of 7 ranked benchmarks measured

Reasoning

Not enough dataFull reasoning ranking

0 of 6 ranked benchmarks measured

Coding

Not enough dataFull coding ranking

0 of 10 ranked benchmarks measured

More results

Benchmarks outside the capability baskets. They are not ranked against a leader.

GPT-5.1 Codex mini: common questions

Who makes GPT-5.1 Codex mini?

GPT-5.1 Codex mini is made by OpenAI.

When was GPT-5.1 Codex mini released?

GPT-5.1 Codex mini was released on Nov 13, 2025, according to Artificial Analysis.

What is GPT-5.1 Codex mini good at?

GPT-5.1 Codex mini is behind the leaders in factuality, long context, instruction following, multimodal tasks, and agentic tasks. Too few results yet to rate reasoning, coding, safety, math, or multilingual tasks.

How much does GPT-5.1 Codex mini cost?

GPT-5.1 Codex mini costs $0.25 per million input tokens and $2.00 per million output tokens, according to Artificial Analysis. We track its price at 2 providers. At a mix of three input tokens to one output token, it is cheaper than 50% of the 330 priced models we track.

How many benchmarks has GPT-5.1 Codex mini been tested on?

We track 18 results for GPT-5.1 Codex mini on 18 benchmarks from 3 sources, 1 of them independently verified. The latest was recorded on Oct 8, 2026.

Which API features does GPT-5.1 Codex mini support?

OpenRouter lists tool calling, structured outputs, and reasoning for GPT-5.1 Codex mini.

About this record

Where GPT-5.1 Codex mini's numbers come from, and every name it appears under.

Tracked since
Apr 27, 2026
Newest source mention
May 2, 2026

Where the results come from

Verification: 18 scores · 1 independently verified · 16 aggregator-attributed · 1 vendor-reported. How these tiers are assigned

From 3 sources on 3 sites. Artificial Analysis supplies 16 of them; the 1 independently verified result comes from 1 site. Bars are coloured by trust tier.

  • artificialanalysis.ai16
  • api.llm-stats.com1
  • datasets-server.huggingface.co1

Also known as

gpt-5-1-codex-minigpt-5.1 codex mini (high)