Models / OpenAI/ GPT-5.6 Sol

GPT-5.6 Sol

OpenAI · released Jul 9, 2026

Input: text, images and documents. Output: text.InputOutput
Type
Closed
Input
$2.00
Output
$10.00
Cached
$0.20

List price · per 1M tokens · OpenAI at 1.1M context · source ↗

Our take

Written Aug 4, 2026

OpenAI's top-tier reasoning and coding model can handle a 1,050,000-token request. It is strongest when your team is already using OpenAI's tooling and can use the cheaper first-party tier.

Who should pick it

Choose this if frontier reasoning quality fits your budget on the base first-party tier, or if you are committed to Azure and need frontier capability inside your tenancy. Skip it if price tiers confuse your procurement, or if you need consistent pricing across every host.

The case for it

  • Base first-party tier is roughly half the input and output price of Claude Opus 5.
  • Slightly larger request limit than the one-million-token peers.

The case against it

  • Pricing spans multiple tiers and hosts. The same model can cost markedly more depending on the endpoint.
  • Not downloadable, so API access only; no self-hosting path.
00

How good is it?

A closed text model for everyday questions, drafting, coding and multi-step agent work.

Good at
  • getting answers to everyday questionsArena Text (overall) · 15th of 168
  • drafts, rewrites and editingArena Creative Writing · 11th of 168
  • writing and completing codeArena Coding · 13th of 168
  • multi-step work it carries out for youArena Agent · 6th of 55
  • calling tools to carry out requestsArena Agent · Tool use · 9th of 55
  • changing course when you give new instructionsArena Agent · Steerability · 9th of 55

EverydayGeneral questions and everyday reasoning

4 of 5

Arena Text (overall)15th of 168 · 1483

Arena Hard Prompts 13th of 168Arena Maths 16th of 163LiveBench Reasoning 4th of 58LiveBench Mathematics 6th of 58LiveBench Data Analysis 7th of 58

CodingWriting and fixing code on its own

4.5 of 5

Arena Coding13th of 168 · 1531

Arena Code (WebDev) 16th of 95LiveBench Coding 5th of 58

AgenticPlanning, calling tools, staying on task

3.5 of 5

Arena Agent6th of 55 · 0.062

LiveBench Agentic Coding 23rd of 58

WritingDrafting and rewriting prose

4 of 5

Arena Creative Writing11th of 168 · 1470

LiveBench Language 6th of 58
How it behaves in an agent loop
Tool usereaches for the right one, and does not invent one9th of 55
Steerabilitydoes what it was asked, and changes course when told9th of 55
Recoverygets back on track after a command fails23rd of 55
Task outcomefinishes what the session set out to do19th of 55

Placings on Arena's agent boards, from live sessions people ran themselves. A model can lead on one of these and sit mid-field on the others.

Other boards it appears on
Arena Instruction Following 7th of 168LiveBench 7th of 58LiveBench Instruction Following 19th of 58Arena Agent · Steerability 9th of 55Arena Agent · Tool use 9th of 55Arena Agent · Task outcome 19th of 55Arena Agent · Recovery 23rd of 55

Boards this model appears on that none of the ratings above are built on.

Every published score for this model20 scoresEvery figure we hold, from 20 boards, with who ran it and a link to the source — including the boards no rating above is built on.
LiveBenchreasoning
81.05source ↗
56.21source ↗
83.94source ↗
79.84source ↗
71.85source ↗
87.68source ↗
96.2source ↗
91.65source ↗
0.062source ↗
0.028source ↗
0.066source ↗
0.029source ↗
0.004source ↗
1531source ↗
1470source ↗
1509source ↗
1489source ↗
1494source ↗
1483source ↗
1617source ↗
01

Where to rent it

Prices checked 4 hours ago — each listing carries its own date.

Cheapest published offer

OpenAI, through OpenRouter

The lab is the cheapest at this context. The strip above and this offer are the same one, compared at 1.1M of context. One cheaper row below is outside that comparison: a non-standard pricing tier.

per 1M tokens
$2.00 in / $10.00 out
Context served
1.1M
Throughput
~41 tok/s
Current provider offers with price, context and prompt-privacy answers
ProviderIn / out per 1M tokensContextThroughputTrains on promptsLogs promptsZero retention
OpenAIflex tierThrough OpenRouter$1.00 / $5.00checked 4 hours ago1.1M128K max reply96 tok/sNoYesunknown periodUnknown
OpenRouterOpenRouter's own listing$2.00 / $10.00checked 4 hours ago1.1Mnot measuredUnknownUnknownUnknown
OpenAIThrough OpenRouter$2.00 / $10.00checked 4 hours ago1.1M128K max reply41 tok/sNoYesunknown periodUnknown
OpenAIfast tierThrough OpenRouter$4.00 / $20.00checked 4 hours ago1.1M128K max reply80 tok/sNoYesunknown periodUnknown
Microsoft Azure AIThrough OpenRouter$4.00 / $20.00checked 4 hours ago1.1M128K max reply25 tok/sNoNoConfirmed
Microsoft Azure AIusThrough OpenRouter$4.40 / $22.00checked 4 hours ago1.1M128K max reply76 tok/sNoNoConfirmed
Amazon Bedrockus-east-1Through OpenRouter$4.40 / $22.00checked 4 hours ago1.1M128K max reply72 tok/sNoNoUnknown
Microsoft Azure AIeuThrough OpenRouter$4.40 / $22.00checked 4 hours ago1.1M128K max reply68 tok/sNoNoConfirmed

Across the 8 listings we hold: 7 say they do not train on prompts, 0 say they do and 1 does not say. 3 appear in the zero-retention registry we check; the rest are unknown to us.

What each host's API supports

From the parameter list each endpoint publishes. Streaming is omitted: nothing we hold reports it, for any model.

API features per host
ProviderTool callingJSON outputStrict schema
OpenAIflexThrough OpenRouter✓✓✓
OpenRouterOpenRouter's own listing✓✓✓
OpenAIThrough OpenRouter✓✓✓
OpenAIfastThrough OpenRouter✓✓✓
Microsoft Azure AIThrough OpenRouter✓✓✓
Microsoft Azure AIusThrough OpenRouter✓✓✓
Amazon Bedrockus-east-1Through OpenRouter✓✗✗
Microsoft Azure AIeuThrough OpenRouter✓✓✓

Tool calling: 8 of 8 listings say yes. JSON output: 7 of 8 listings say yes, 1 says no. Strict schema: 7 of 8 listings say yes, 1 says no.

02

Models people weigh against GPT-5.6 Sol

03

When we formed this view

Recent changes

Sep 25, 2026BenchmarkScored 0.062 via xHigh on Arena Agent
What movedleaderboard
Sep 25, 2026BenchmarkScored 0.028 via xHigh on Arena Agent · Recovery
What movedleaderboard
Sep 25, 2026BenchmarkScored 0.066 via xHigh on Arena Agent · Steerability
What movedleaderboard
Sep 25, 2026BenchmarkScored 0.029 via xHigh on Arena Agent · Task outcome
What movedleaderboard
Sep 25, 2026BenchmarkScored 0.004 via xHigh on Arena Agent · Tool use
What movedleaderboard
Sep 25, 2026BenchmarkScored 1531 via xHigh on Arena Coding
What movedleaderboard
Sep 25, 2026BenchmarkScored 1470 via xHigh on Arena Creative Writing
What movedleaderboard
Sep 25, 2026BenchmarkScored 1509 via xHigh on Arena Hard Prompts
What movedleaderboard
Sep 25, 2026BenchmarkScored 1489 via xHigh on Arena Instruction Following
What movedleaderboard
Sep 25, 2026BenchmarkScored 1494 via xHigh on Arena Maths
What movedleaderboard

Each date is the day we first saw the change, or the day the maker announced it.

What we do not know about this model yet

  • Nothing we hold says whether an endpoint streams, so we do not show it either way.
  • 1 of 8 listings does not say whether it trains on prompts.
  • We hold no batch or off-peak rate for any of its listings.
04

Licence and identifiers

What the licence allowsWe hold no licence record for this model. Inside are the identifiers you need to pull it — its Hugging Face repo where we have one, our slug and a machine-readable card.

Licence

We hold no licence record for this model, and no record of published weights either — so we can neither summarise its terms nor point you at the weights.

Identifiers

Takes in, gives back
Text, images and documents in, text out
Catalogue slug
openai-gpt-5-6-sol

Machine-readable model card (omc.json) →

Something wrong on this page? Tell us