GPT-5 mini
GPT-5 mini is capable in instruction following and long context; and behind the leaders in multimodal tasks, coding, agentic tasks, and math. Too few results yet to rate reasoning, safety, multilingual tasks, or factuality.
Capability profile
Bars show the model's median result as a share of the leading model's, per capability. Select a row to see its results.
Too few results yet to rate Reasoning, Safety, Multilingual or Factuality.
Price
$0.25input$2.00outputper million tokens
From Artificial Analysis · 2 providers tracked · All prices
Evidence
96results on46benchmarks
- 37 independently verified
- 49 aggregator
- 5 vendor-reported
- 5 cross-referenced
From 22 sources · latest Oct 8, 2026 · How verification works
API features
Tool callingStructured outputsReasoning
As listed by OpenRouter
Research
57 papers reference GPT-5 miniGPT-5 mini benchmark results
96 results on 46 benchmarks, grouped by capability. Each bar is the result as a share of the capability leader's; every result links to its source.
21.9% behind the leader2 of 3 ranked benchmarks measured
- IFBench75.44Oct 8, 2026aa_ifbench
- 58.99Oct 8, 2026
22.5% behind the leader1 of 3 ranked benchmarks measured
- 72.33Oct 8, 2026
Show 1 more long context resultHide 1 long context result
- 39.33Oct 8, 2026
31.9% behind the leader1 of 6 ranked benchmarks measured
- MMMU-Pro70.06Oct 8, 2026aa_mmmu_pro
Show 3 more multimodal resultsHide 3 multimodal results
39.8% behind the leader4 of 10 ranked benchmarks measured
- 59.80Sep 25, 2026
- SciCode39.00Oct 8, 2026aa_scicode
- Terminal-Bench Hard33.33Oct 8, 2026aa_terminalbench_hard
- Terminal-Bench 2.13.75Oct 8, 2026terminalbenchV21
Show 6 more coding resultsHide 6 coding results
- SciCode40.97Sep 4, 2026aa_scicode
- SciCode36.92Sep 4, 2026aa_scicode
- 39.70May 1, 2026
- 56.20May 1, 2026
- Terminal-Bench Hard28.79Oct 8, 2026aa_terminalbench_hard
- Terminal-Bench Hard14.39Oct 8, 2026aa_terminalbench_hard
43.8% behind the leader3 of 7 ranked benchmarks measured
- τ-Bench V3 · Banking15.46Oct 8, 2026tauBanking
- 13.48Oct 8, 2026
- Terminal-Bench 4.00.00Oct 8, 2026
Show 2 more agentic resultsHide 2 agentic results
- 25.07Jun 15, 2026
- 0.00Jun 15, 2026
47.1% behind the leader2 of 5 ranked benchmarks measured
- 46.67Sep 21, 2026
- 12.20Sep 21, 2026
Show 2 more math resultsHide 2 math results
- 5.96Sep 22, 2026
- 18.25Sep 21, 2026
0 of 6 ranked benchmarks measured
- 1.67Sep 22, 2026
- 4.44May 15, 2026
- 0.83May 10, 2026
Show 13 more reasoning resultsHide 13 reasoning results
- 4.03May 10, 2026
- 1.43Oct 8, 2026
- 0.00Oct 8, 2026
- GPQA Diamond80.30Oct 8, 2026gpqa
- GPQA Diamond68.69Oct 8, 2026gpqa
- GPQA Diamond82.83Oct 8, 2026gpqa
- GPQA Diamond82.30Oct 7, 2026GPQA
- 82.30Jul 13, 2026
- Humanity's Last Exam15.94Oct 8, 2026aa_hle
- Humanity's Last Exam5.05Oct 8, 2026aa_hle
- Humanity's Last Exam21.46Oct 8, 2026aa_hle
- 16.70Oct 7, 2026
- Humanity's Last Exam16.70Jul 13, 2026HLE (no tools)
0 of 4 ranked benchmarks measured
- 21.60May 20, 2026
- 12.90May 2, 2026
- AA-Omniscience · Accuracy19.17Oct 8, 2026omniscienceAccuracy
Show 5 more factuality resultsHide 5 factuality results
- AA-Omniscience · Accuracy23.07Oct 8, 2026omniscienceAccuracy
- AA-Omniscience · Accuracy24.97Oct 8, 2026omniscienceAccuracy
- AA-Omniscience · Non-hallucination55.68Oct 8, 2026omniscienceNonHallucination
- AA-Omniscience · Non-hallucination10.66Oct 8, 2026omniscienceNonHallucination
- AA-Omniscience · Non-hallucination43.65Oct 8, 2026omniscienceNonHallucination
More results
Benchmarks outside the capability baskets. They are not ranked against a leader.
- 5.33Sep 22, 2026
- AA Intelligence17.00Sep 21, 2026Artificial Analysis Intelligence Index
- 97.67Sep 7, 2026
- 100.00Sep 7, 2026
- 97.06Sep 7, 2026
- 99.11Sep 7, 2026
Show 37 more resultsHide 37 results
- 8.92Sep 9, 2026
- 40.92Jun 18, 2026
- 13.41Jun 18, 2026
- AA Intelligence26.00Jun 20, 2026Artificial Analysis Intelligence Index
- AA Intelligence9.91Oct 8, 2026aa_intelligence_index
- AA Intelligence20.60Oct 8, 2026aa_intelligence_index
- AA Intelligence16.77Oct 8, 2026aa_intelligence_index
- -11.03Oct 8, 2026
- -53.05Oct 8, 2026
- -17.32Oct 8, 2026
- 87.50May 20, 2026
- 91.10Oct 7, 2026
- AIME 202591.10Jul 13, 2026AIME 2025 (no tools)
- 85.71May 19, 2026
- 54.33May 15, 2026
- 26.33May 10, 2026
- 37.33May 10, 2026
- Artificial Analysis Coding Index15.58Sep 9, 2026aa_coding_index
- 32.85Jun 18, 2026
- Artificial Analysis Coding Index21.90Jun 18, 2026aa_coding_index
- 96.30Sep 7, 2026
- FrontierMath (overall)22.10Oct 7, 2026FrontierMath
- frontiermath_tier_4_v16.25Aug 29, 2026frontiermath_tier_4
- frontiermath_tier_4_v14.17May 20, 2026frontiermath_tier_4
- 87.80Oct 7, 2026
- HMMT 202587.80Jul 13, 2026HMMT 2025 (no tools)
- 84.17May 10, 2026
- 77.40Jul 13, 2026
- 78.16Sep 2, 2026
- 59.80Aug 29, 2026
- 56.20May 1, 2026
- vectara_answer_rate99.90May 2, 2026Answer Rate
- vectara_avg_summary_length169.70May 2, 2026Average Summary Length (Words)
- vectara_factual_consistency87.10May 2, 2026Factual Consistency Rate
- τ²-Bench Telecom (AA run)71.05Oct 8, 2026aa_tau2
- τ²-Bench Telecom (AA run)31.87Oct 8, 2026aa_tau2
- τ²-Bench Telecom (AA run)68.42Oct 8, 2026aa_tau2
GPT-5 mini: common questions
Who makes GPT-5 mini?
GPT-5 mini is made by OpenAI.
When was GPT-5 mini released?
GPT-5 mini was released on Aug 7, 2025, according to Artificial Analysis.
What is GPT-5 mini good at?
GPT-5 mini is capable in instruction following and long context; and behind the leaders in multimodal tasks, coding, agentic tasks, and math. Too few results yet to rate reasoning, safety, multilingual tasks, or factuality.
How much does GPT-5 mini cost?
GPT-5 mini costs $0.25 per million input tokens and $2.00 per million output tokens, according to Artificial Analysis. We track its price at 2 providers. At a mix of three input tokens to one output token, it is cheaper than 50% of the 331 priced models we track.
How many benchmarks has GPT-5 mini been tested on?
We track 96 results for GPT-5 mini on 46 benchmarks from 22 sources, 37 of them independently verified. The latest was recorded on Oct 8, 2026.
Which API features does GPT-5 mini support?
OpenRouter lists tool calling, structured outputs, and reasoning for GPT-5 mini.
About this record
Where GPT-5 mini's numbers come from, and every name it appears under.
- Tracked since
- May 10, 2026
- Newest source mention
- Aug 25, 2026
Where the results come from
Verification: 96 scores · 37 independently verified · 49 aggregator-attributed · 5 vendor cross-reference · 5 vendor-reported. How these tiers are assigned
From 22 sources on 11 sites. Artificial Analysis supplies 51 of them; the 37 independently verified results come from 9 sites. Bars are coloured by trust tier.
- artificialanalysis.ai51
- arcprize.org8
- epoch.ai7
- storage.googleapis.com6
- api.llm-stats.com5
- swebench.com5
- x.ai5
- raw.githubusercontent.com4
- matharena.ai3
- datasets-server.huggingface.co1
- labs.scale.com1
Also known as
How our sources name GPT-5 mini at each reasoning setting.
| Setting | Short form | API id |
|---|---|---|
| minimal | gpt-5 mini (minimal) | gpt-5-mini-minimal |
| low | gpt-5 mini (low) | — |
| medium | gpt-5 mini (2025-08-07) (medium reasoning) gpt-5 mini (medium) gpt 5 mini (2025-08-07) (medium) | gpt-5-mini-medium |
| high | gpt-5 mini (high) | GPT-5-mini (high) (Non-Reasoning) gpt-5-mini (high) (reasoning) gpt-5-mini-high |