GPT-5.1 Codex mini
GPT-5.1 Codex mini is behind the leaders in factuality, long context, instruction following, multimodal tasks, and agentic tasks. Too few results yet to rate reasoning, coding, safety, math, or multilingual tasks.
Capability profile
Bars show the model's median result as a share of the leading model's, per capability. Select a row to see its results.
Too few results yet to rate Reasoning, Coding, Safety, Math or Multilingual.
Price
$0.25input$2.00outputper million tokens
From Artificial Analysis · 2 providers tracked · All prices
Evidence
18results on18benchmarks
- 1 independently verified
- 16 aggregator
- 1 vendor-reported
From 3 sources · latest Oct 8, 2026 · How verification works
API features
Tool callingStructured outputsReasoning
As listed by OpenRouter
GPT-5.1 Codex mini benchmark results
18 results on 18 benchmarks, grouped by capability. Each bar is the result as a share of the capability leader's; every result links to its source.
27.6% behind the leader2 of 4 ranked benchmarks measured
- AA-Omniscience · Non-hallucination48.42Oct 8, 2026omniscienceNonHallucination
- AA-Omniscience · Accuracy23.20Oct 8, 2026omniscienceAccuracy
27.7% behind the leader1 of 3 ranked benchmarks measured
- 66.67Oct 8, 2026
30.2% behind the leader1 of 3 ranked benchmarks measured
- IFBench67.89Oct 8, 2026aa_ifbench
32.8% behind the leader1 of 6 ranked benchmarks measured
- MMMU-Pro69.02Oct 8, 2026aa_mmmu_pro
37.7% behind the leader1 of 7 ranked benchmarks measured
- 27.70Jun 15, 2026
0 of 6 ranked benchmarks measured
- 0.00Oct 8, 2026
- GPQA Diamond81.31Oct 8, 2026gpqa
- Humanity's Last Exam18.49Oct 8, 2026aa_hle
0 of 10 ranked benchmarks measured
- 1243.35May 22, 2026
- Terminal-Bench Hard33.33Oct 8, 2026aa_terminalbench_hard
- SciCode42.59Sep 4, 2026aa_scicode
More results
Benchmarks outside the capability baskets. They are not ranked against a leader.
- τ²-Bench Telecom (AA run)62.87Oct 8, 2026aa_tau2
- AA Intelligence20.38Oct 8, 2026aa_intelligence_index
- -16.42Oct 8, 2026
- 36.42Jun 18, 2026
- 38.67Jun 18, 2026
- 42.10Oct 7, 2026
GPT-5.1 Codex mini: common questions
Who makes GPT-5.1 Codex mini?
GPT-5.1 Codex mini is made by OpenAI.
When was GPT-5.1 Codex mini released?
GPT-5.1 Codex mini was released on Nov 13, 2025, according to Artificial Analysis.
What is GPT-5.1 Codex mini good at?
GPT-5.1 Codex mini is behind the leaders in factuality, long context, instruction following, multimodal tasks, and agentic tasks. Too few results yet to rate reasoning, coding, safety, math, or multilingual tasks.
How much does GPT-5.1 Codex mini cost?
GPT-5.1 Codex mini costs $0.25 per million input tokens and $2.00 per million output tokens, according to Artificial Analysis. We track its price at 2 providers. At a mix of three input tokens to one output token, it is cheaper than 50% of the 330 priced models we track.
How many benchmarks has GPT-5.1 Codex mini been tested on?
We track 18 results for GPT-5.1 Codex mini on 18 benchmarks from 3 sources, 1 of them independently verified. The latest was recorded on Oct 8, 2026.
Which API features does GPT-5.1 Codex mini support?
OpenRouter lists tool calling, structured outputs, and reasoning for GPT-5.1 Codex mini.
About this record
Where GPT-5.1 Codex mini's numbers come from, and every name it appears under.
- Tracked since
- Apr 27, 2026
- Newest source mention
- May 2, 2026
Where the results come from
Verification: 18 scores · 1 independently verified · 16 aggregator-attributed · 1 vendor-reported. How these tiers are assigned
From 3 sources on 3 sites. Artificial Analysis supplies 16 of them; the 1 independently verified result comes from 1 site. Bars are coloured by trust tier.
- artificialanalysis.ai16
- api.llm-stats.com1
- datasets-server.huggingface.co1