Skip to content
VECTOR WIREAI INTELLIGENCE
UTC

GPT-5.6 Sol

Score basis

GPT-5.6 Sol is at the frontier in agentic tasks; strong in reasoning, long context, coding, and math; and capable in multimodal tasks, factuality, and instruction following. Too few results yet to rate safety or multilingual tasks.

Capability profile

Bars show the model's median result as a share of the leading model's, per capability. Select a row to see its results.

Frontier1 capability
  1. Agentic−4.2%7 of 7
Strong4 capabilities
  1. Reasoning−2.6%6 of 6
  2. Long Context−5.1%1 of 3
  3. Coding−6.2%7 of 10
  4. Math−9.3%5 of 5
Capable3 capabilities
  1. Multimodal−18.3%2 of 6
  2. Factuality−18.7%3 of 4
  3. Instruction Following−21.0%2 of 3
Not rated2 capabilities

Too few results yet to rate Safety or Multilingual.

Price

$4.00input$20.00outputper million tokens

From Artificial Analysis · 2 providers tracked · All prices

Costs more than 92% of 330 priced models · 3:1 input-to-output blend, log scale

Evidence

199results on119benchmarks

  • 28 independently verified
  • 70 aggregator
  • 27 vendor-reported
  • 74 cross-referenced

From 33 sources · latest Oct 8, 2026 · How verification works

API features

Tool callingStructured outputsReasoning

As listed by OpenRouter

GPT-5.6 Sol benchmark results

199 results on 119 benchmarks, grouped by capability. Each bar is the result as a share of the capability leader's; every result links to its source.

Agentic

FrontierFull agentic ranking

4.2% behind the leader7 of 7 ranked benchmarks measured

Show 15 more agentic resultsHide 15 agentic results

Reasoning

StrongFull reasoning ranking

2.6% behind the leader6 of 6 ranked benchmarks measured

Show 14 more reasoning resultsHide 14 reasoning results

Long Context

StrongFull long context ranking

5.1% behind the leader1 of 3 ranked benchmarks measured

Show 4 more long context resultsHide 4 long context results

6.2% behind the leader7 of 10 ranked benchmarks measured

Show 8 more coding resultsHide 8 coding results

9.3% behind the leader5 of 5 ranked benchmarks measured

Show 2 more math resultsHide 2 math results

Multimodal

CapableFull multimodal ranking

18.3% behind the leader2 of 6 ranked benchmarks measured

Show 5 more multimodal resultsHide 5 multimodal results

Factuality

CapableFull factuality ranking

18.7% behind the leader3 of 4 ranked benchmarks measured

Show 8 more factuality resultsHide 8 factuality results

Instruction Following

CapableFull instruction following ranking

21.0% behind the leader2 of 3 ranked benchmarks measured

Show 2 more instruction following resultsHide 2 instruction following results

More results

Benchmarks outside the capability baskets. They are not ranked against a leader.

Show 105 more resultsHide 105 results

GPT-5.6 Sol: common questions

Who makes GPT-5.6 Sol?

GPT-5.6 Sol is made by OpenAI.

When was GPT-5.6 Sol released?

GPT-5.6 Sol was released on Jul 9, 2026, according to Artificial Analysis.

What is GPT-5.6 Sol good at?

GPT-5.6 Sol is at the frontier in agentic tasks; strong in reasoning, long context, coding, and math; and capable in multimodal tasks, factuality, and instruction following. Too few results yet to rate safety or multilingual tasks.

How much does GPT-5.6 Sol cost?

GPT-5.6 Sol costs $4.00 per million input tokens and $20.00 per million output tokens, according to Artificial Analysis. We track its price at 2 providers. At a mix of three input tokens to one output token, it costs more than 92% of the 330 priced models we track.

How many benchmarks has GPT-5.6 Sol been tested on?

We track 199 results for GPT-5.6 Sol on 119 benchmarks from 33 sources, 28 of them independently verified. The latest was recorded on Oct 8, 2026.

Which API features does GPT-5.6 Sol support?

OpenRouter lists tool calling, structured outputs, and reasoning for GPT-5.6 Sol.

About this record

Where GPT-5.6 Sol's numbers come from, and every name it appears under.

Tracked since
Jul 1, 2026
Newest source mention
Oct 7, 2026

Where the results come from

Verification: 199 scores · 28 independently verified · 70 aggregator-attributed · 74 vendor cross-reference · 27 vendor-reported. How these tiers are assigned

From 33 sources on 16 sites. Artificial Analysis supplies 70 of them; the 28 independently verified results come from 8 sites. Bars are coloured by trust tier.

  • artificialanalysis.ai70
  • huggingface.co60
  • api.llm-stats.com15
  • deploymentsafety.openai.com12
  • arcprize.org8
  • livebench.ai7
  • deepmind.google6
  • anthropic.com4
  • raw.githubusercontent.com4
  • epoch.ai3
  • www-cdn.anthropic.com3
  • datasets-server.huggingface.co2
  • labs.scale.com2
  • lmarena.ai1
  • simple-bench.com1
  • x.ai1

Also known as

How our sources name GPT-5.6 Sol at each reasoning setting.

SettingShort formAPI id
lowgpt-5.6 sol (low) gpt 5.6 sol lowgpt-5-6-sol-low
mediumgpt-5.6 sol (medium)gpt-5-6-sol-medium
highgpt-5.6 sol (high) gpt-5.6 sol highgpt-5-6-sol-high
xhighgpt-5.6 sol (xhigh)gpt-5.6-sol-xhigh gpt-5.6-sol-xhigh (codex-harness) gpt-5-6-sol-xhigh
maxgpt-5.6 sol (max) gpt-5.6 sol max gpt sol 5.6 max gpt-5.6 sol (pro, max)gpt-5.6-sol (max, via codex)
Also listed asgpt-5.6 sol (non-reasoning)gpt-sol 5.6gpt-5-6-sol-non-reasoninggpt-5-6-solgpt-5.6 (sol)openai.gpt-5.6-solopenai's gpt-5.6-solgpt-5.6-sol:batch5.6 sol5.6 soulopenai 5.6 solcodex/gpt-5.6 solgpt‑5.6 solgpt-5.6sol