GPT-4.1 nano
GPT-4.1 nano is behind the leaders in multimodal tasks. Too few results yet to rate reasoning, coding, agentic tasks, safety, long context, math, multilingual tasks, instruction following, or factuality.
Capability profile
Bars show the model's median result as a share of the leading model's, per capability. Select a row to see its results.
Too few results yet to rate Reasoning, Coding, Agentic, Safety, Long Context, Math, Multilingual, Instruction Following or Factuality.
Price
$0.10input$0.40outputper million tokens
From Artificial Analysis · 2 providers tracked · All prices
Evidence
39results on36benchmarks
- 11 independently verified
- 18 aggregator
- 10 vendor-reported
From 13 sources · latest Oct 8, 2026 · How verification works
API features
Tool callingStructured outputs
As listed by OpenRouter
Research
6 papers reference GPT-4.1 nanoGPT-4.1 nano benchmark results
39 results on 36 benchmarks, grouped by capability. Each bar is the result as a share of the capability leader's; every result links to its source.
61.1% behind the leader4 of 6 ranked benchmarks measured
- 55.40Oct 7, 2026
- 56.20Oct 7, 2026
- MMMU-Pro40.12Oct 8, 2026aa_mmmu_pro
- CharXiv (reasoning)40.50Oct 7, 2026CharXiv-R
0 of 6 ranked benchmarks measured
- 0.00May 10, 2026
- 0.00Oct 8, 2026
- GPQA Diamond51.21Oct 8, 2026gpqa
Show 2 more reasoning resultsHide 2 reasoning results
- GPQA Diamond50.30Oct 7, 2026GPQA
- Humanity's Last Exam3.75Oct 8, 2026aa_hle
0 of 10 ranked benchmarks measured
- Terminal-Bench 2.13.75Oct 8, 2026terminalbenchV21
- Terminal-Bench Hard3.79Oct 8, 2026aa_terminalbench_hard
- SciCode25.93Sep 4, 2026aa_scicode
0 of 7 ranked benchmarks measured
- 0.00Oct 8, 2026
- τ-Bench V3 · Banking3.51Oct 8, 2026tauBanking
0 of 3 ranked benchmarks measured
- 20.33Oct 8, 2026
0 of 3 ranked benchmarks measured
- IFBench32.04Oct 8, 2026aa_ifbench
- 15.00Oct 7, 2026
0 of 4 ranked benchmarks measured
- 6.00Sep 1, 2026
- AA-Omniscience · Accuracy13.68Oct 8, 2026omniscienceAccuracy
- AA-Omniscience · Non-hallucination17.44Oct 8, 2026omniscienceNonHallucination
More results
Benchmarks outside the capability baskets. They are not ranked against a leader.
- AA Intelligence8.00Aug 7, 2026Artificial Analysis Intelligence Index
- 8.90Jun 19, 2026
- 61.53May 19, 2026
- 87.50May 10, 2026
- 96.00May 10, 2026
- 99.00May 10, 2026
Show 13 more resultsHide 13 results
- 1.17Aug 18, 2026
- AA Intelligence7.82Oct 8, 2026aa_intelligence_index
- -57.58Oct 8, 2026
- 9.80Oct 7, 2026
- 99.58May 10, 2026
- 0.00May 10, 2026
- Artificial Analysis Coding Index11.14Sep 9, 2026aa_coding_index
- 86.75May 10, 2026
- 74.50Aug 31, 2026
- 66.90Oct 7, 2026
- TAU-bench (airline)14.00Oct 7, 2026TAU-bench Airline
- TAU-bench (retail)22.60Oct 7, 2026TAU-bench Retail
- τ²-Bench Telecom (AA run)17.25Oct 8, 2026aa_tau2
GPT-4.1 nano: common questions
Who makes GPT-4.1 nano?
GPT-4.1 nano is made by OpenAI.
When was GPT-4.1 nano released?
GPT-4.1 nano was released on Apr 14, 2025, according to Artificial Analysis.
What is GPT-4.1 nano good at?
GPT-4.1 nano is behind the leaders in multimodal tasks. Too few results yet to rate reasoning, coding, agentic tasks, safety, long context, math, multilingual tasks, instruction following, or factuality.
How much does GPT-4.1 nano cost?
GPT-4.1 nano costs $0.10 per million input tokens and $0.40 per million output tokens, according to Artificial Analysis. We track its price at 2 providers. At a mix of three input tokens to one output token, it is cheaper than 82% of the 330 priced models we track.
How many benchmarks has GPT-4.1 nano been tested on?
We track 39 results for GPT-4.1 nano on 36 benchmarks from 13 sources, 11 of them independently verified. The latest was recorded on Oct 8, 2026.
Which API features does GPT-4.1 nano support?
OpenRouter lists tool calling and structured outputs for GPT-4.1 nano.
About this record
Where GPT-4.1 nano's numbers come from, and every name it appears under.
- Tracked since
- Apr 27, 2026
- Newest source mention
- Aug 9, 2026
Where the results come from
Verification: 39 scores · 11 independently verified · 18 aggregator-attributed · 10 vendor-reported. How these tiers are assigned
From 13 sources on 6 sites. Artificial Analysis supplies 19 of them; the 11 independently verified results come from 5 sites. Bars are coloured by trust tier.
- artificialanalysis.ai19
- api.llm-stats.com10
- storage.googleapis.com6
- arcprize.org2
- aider.chat1
- epoch.ai1