GPT-5.4 nano
GPT-5.4 nano is capable in long context and instruction following; and behind the leaders in reasoning, multimodal tasks, coding, factuality, agentic tasks, and math. Too few results yet to rate safety or multilingual tasks.
Capability profile
Bars show the model's median result as a share of the leading model's, per capability. Select a row to see its results.
Too few results yet to rate Safety or Multilingual.
Price
$0.20input$1.25outputper million tokens
From Artificial Analysis · 2 providers tracked · All prices
Evidence
92results on47benchmarks
- 27 independently verified
- 54 aggregator
- 11 vendor-reported
From 11 sources · latest Oct 8, 2026 · How verification works
API features
Tool callingStructured outputsReasoning
As listed by OpenRouter
Research
9 papers reference GPT-5.4 nanoGPT-5.4 nano benchmark results
92 results on 47 benchmarks, grouped by capability. Each bar is the result as a share of the capability leader's; every result links to its source.
16.7% behind the leader1 of 3 ranked benchmarks measured
- 76.67Oct 8, 2026
23.8% behind the leader2 of 3 ranked benchmarks measured
- IFBench75.92Oct 8, 2026aa_ifbench
- LiveBench · Instruction Following67.20Oct 8, 2026livebench_instruction_following@2026-06-25
29.4% behind the leader5 of 6 ranked benchmarks measured
- LiveBench · Reasoning81.10Oct 8, 2026livebench_reasoning@2026-06-25
- GPQA Diamond81.72Oct 8, 2026gpqa
- Humanity's Last Exam28.27Oct 8, 2026aa_hle
- 9.25Oct 8, 2026
- 5.69May 10, 2026
Show 11 more reasoning resultsHide 11 reasoning results
- 1.53Sep 21, 2026
- 1.94Sep 21, 2026
- 3.61May 10, 2026
- 0.00Oct 8, 2026
- 5.14Oct 8, 2026
- GPQA Diamond76.06Oct 8, 2026gpqa
- GPQA Diamond55.76Oct 8, 2026gpqa
- GPQA Diamond82.80Oct 7, 2026GPQA
- Humanity's Last Exam15.94Oct 8, 2026aa_hle
- Humanity's Last Exam4.08Oct 8, 2026aa_hle
- 24.30Oct 7, 2026
33.9% behind the leader1 of 6 ranked benchmarks measured
- MMMU-Pro65.38Oct 8, 2026aa_mmmu_pro
Show 5 more multimodal resultsHide 5 multimodal results
35.1% behind the leader6 of 10 ranked benchmarks measured
- LiveBench · Coding70.84Oct 8, 2026livebench_coding@2026-06-25
- SciCode47.22Oct 8, 2026aa_scicode
- Terminal-Bench 2.160.67Oct 8, 2026terminalbenchV21
- Terminal-Bench Hard42.42Oct 8, 2026aa_terminalbench_hard
- LiveBench · Agentic Coding46.77Oct 8, 2026livebench_agentic_coding@2026-06-25
- 52.40Oct 7, 2026
Show 4 more coding resultsHide 4 coding results
- SciCode35.19Sep 4, 2026aa_scicode
- SciCode38.43Sep 4, 2026aa_scicode
- Terminal-Bench Hard33.33Oct 8, 2026aa_terminalbench_hard
- Terminal-Bench Hard24.24Oct 8, 2026aa_terminalbench_hard
36.8% behind the leader4 of 4 ranked benchmarks measured
- 3.10May 2, 2026
- AA-Omniscience · Accuracy25.67Oct 8, 2026omniscienceAccuracy
- AA-Omniscience · Non-hallucination25.83Oct 8, 2026omniscienceNonHallucination
- 11.70May 20, 2026
Show 4 more factuality resultsHide 4 factuality results
- AA-Omniscience · Accuracy14.88Oct 8, 2026omniscienceAccuracy
- AA-Omniscience · Accuracy21.90Oct 8, 2026omniscienceAccuracy
- AA-Omniscience · Non-hallucination48.87Oct 8, 2026omniscienceNonHallucination
- AA-Omniscience · Non-hallucination38.48Oct 8, 2026omniscienceNonHallucination
41.5% behind the leader5 of 7 ranked benchmarks measured
- 56.10Oct 7, 2026
- τ-Bench V3 · Banking27.42Oct 8, 2026tauBanking
- 39.00Oct 7, 2026
- 22.53Oct 8, 2026
- Terminal-Bench 4.00.51Oct 8, 2026
Show 5 more agentic resultsHide 5 agentic results
- 24.93Oct 8, 2026
- 24.39Oct 8, 2026
- 12.91Oct 8, 2026
- 9.04Sep 19, 2026
- 34.91Jun 15, 2026
47.2% behind the leader3 of 5 ranked benchmarks measured
- LiveBench · Mathematics90.98Oct 8, 2026livebench_math@2026-06-25
- 44.91Sep 21, 2026
- 12.20Sep 21, 2026
Show 2 more math resultsHide 2 math results
- 20.35Sep 21, 2026
- 4.56Sep 21, 2026
More results
Benchmarks outside the capability baskets. They are not ranked against a leader.
- livebench_data_analysis67.64Oct 8, 2026livebench_data_analysis@2026-06-25
- livebench_language62.51Oct 8, 2026livebench_language@2026-06-25
- 18.33Sep 21, 2026
- 33.00Sep 21, 2026
- frontiermath_tier_4_v16.25Aug 29, 2026frontiermath_tier_4
- 38.17May 10, 2026
Show 24 more resultsHide 24 results
- 17.74Sep 9, 2026
- 25.92Jun 18, 2026
- 41.64Jun 18, 2026
- AA Intelligence20.72Oct 8, 2026aa_intelligence_index
- AA Intelligence20.01Oct 8, 2026aa_intelligence_index
- AA Intelligence11.69Oct 8, 2026aa_intelligence_index
- -29.47Oct 8, 2026
- -37.48Oct 8, 2026
- -18.03Oct 8, 2026
- 51.50May 10, 2026
- Artificial Analysis Coding Index56.07Sep 9, 2026aa_coding_index
- Artificial Analysis Coding Index27.89Jun 18, 2026aa_coding_index
- 35.03Jun 18, 2026
- 38.22Oct 7, 2026
- 70.13Oct 7, 2026
- 75.81Oct 7, 2026
- 46.30Oct 7, 2026
- 35.50Oct 7, 2026
- vectara_answer_rate100.00May 2, 2026Answer Rate
- vectara_avg_summary_length144.40May 2, 2026Average Summary Length (Words)
- vectara_factual_consistency96.90May 2, 2026Factual Consistency Rate
- τ²-Bench Telecom (AA run)76.02Oct 8, 2026aa_tau2
- τ²-Bench Telecom (AA run)52.63Oct 8, 2026aa_tau2
- τ²-Bench Telecom (AA run)34.80Oct 8, 2026aa_tau2
GPT-5.4 nano: common questions
Who makes GPT-5.4 nano?
GPT-5.4 nano is made by OpenAI.
When was GPT-5.4 nano released?
GPT-5.4 nano was released on Mar 17, 2026, according to Artificial Analysis.
What is GPT-5.4 nano good at?
GPT-5.4 nano is capable in long context and instruction following; and behind the leaders in reasoning, multimodal tasks, coding, factuality, agentic tasks, and math. Too few results yet to rate safety or multilingual tasks.
How much does GPT-5.4 nano cost?
GPT-5.4 nano costs $0.20 per million input tokens and $1.25 per million output tokens, according to Artificial Analysis. We track its price at 2 providers. At a mix of three input tokens to one output token, it is cheaper than 60% of the 330 priced models we track.
How many benchmarks has GPT-5.4 nano been tested on?
We track 92 results for GPT-5.4 nano on 47 benchmarks from 11 sources, 27 of them independently verified. The latest was recorded on Oct 8, 2026.
Which API features does GPT-5.4 nano support?
OpenRouter lists tool calling, structured outputs, and reasoning for GPT-5.4 nano.
About this record
Where GPT-5.4 nano's numbers come from, and every name it appears under.
- Tracked since
- Apr 27, 2026
- Newest source mention
- Sep 9, 2026
Where the results come from
Verification: 92 scores · 27 independently verified · 54 aggregator-attributed · 11 vendor-reported. How these tiers are assigned
From 11 sources on 8 sites. Artificial Analysis supplies 54 of them; the 27 independently verified results come from 6 sites. Bars are coloured by trust tier.
- artificialanalysis.ai54
- api.llm-stats.com11
- arcprize.org8
- livebench.ai7
- epoch.ai6
- raw.githubusercontent.com4
- datasets-server.huggingface.co1
- lmarena.ai1
Also known as
How our sources name GPT-5.4 nano at each reasoning setting.
| Setting | Short form | API id |
|---|---|---|
| low | gpt-5.4 nano (low) | — |
| medium | gpt-5.4 nano (medium) | gpt-5-4-nano-medium |
| high | gpt-5.4 nano (high) | gpt-5.4-nano-high |
| xhigh | gpt-5.4 nano (xhigh) | — |