Last verified
FRONTIER REASONINGGPT-5.6 FAMILY1.05M CONTEXTTEXT + VISIONPROMPT CACHING

GPT-5.6 Sol API Pricing

GPT-5.6 Sol is OpenAI's highest-accuracy GPT-5.6 tier for hard reasoning, coding and senior-agent workloads, and it has now had a cut of its own: $4/M input, $20/M output and $0.40/M cached input, against $5 / $30 / $0.50 before. Input-side rates fell to 0.8x and output to two thirds, taking the effective agentic rate from $3.81 to $2.73/M. OpenAI labels the new card promotional and guarantees it at least through November 21, 2026 without publishing what comes after. Sol sat out the July 30 cut that took Luna down 80% and Terra 20% - this one is its own, and it is the only GPT-5 row that moved: Terra, Luna, 5.5, 5.4 and both -pro tiers read unchanged on the same page.

Input - per 1M tokens
$4.00/M
Short context up to 272K input standard
Output - per 1M tokens
$20.00/M
Long context output is $30/M standard
Cached input - 90% off
$0.40/M
Prompt cache read price -90%
Effective - agentic blend
$2.73/M
92/8 split - 82% cache
§ 01 / TERMINAL

Run the numbers.

Live calculator pre-loaded with current GPT-5.6 Sol standard short-context rates. Tweak spend, output mix, or cache hit rate; share the URL to share the calculation.

$ /mo
Workload split
Prompt cache hit rate
Tokens you can process
Words equivalent (English)
Effective rate
Open full calculator (all models · share URL · CSV) →
§ 02 / SCENARIOS

Real-world presets.

§ 03 / TOKENIZER

Paste text. See tokens. See cost.

Exact · o200k_base Auto-counts as you type

Counted with the o200k_base BPE encoding — the same tokenizer the API uses — entirely in your browser. The encoder loads on your first keystroke (~1 MB, one time); your text never leaves this page.

Characters 577
Words 86
Tokens (exact) 151 tokens
Cost as input · uncached $0.0006 USD
Cost as output · uncached $0.00302 USD
Cost as cached input $0.00006 USD
§ 04 / SHELF

Up against the shelf.

All models →
Model Input /M Output /M Effective blended Context Best for
GPT-5.6 Sol Current $4.00 cache $0.40 $20.00 $2.73 agentic 92/8 1.05M Highest-accuracy GPT-5.6 preview
GPT-5.6 Terra $2.00 cache $0.20 $12.00 $1.52 same blend 1.05M Balanced GPT-5.6 preview agents
GPT-5.6 Luna $0.20 cache $0.02 $1.20 $0.152 same blend 1.05M Cheapest GPT-5.6 preview routing
GPT-5.5 $5.00 cache $0.50 $30.00 $3.61 same blend 1M Current broadly available OpenAI flagship
GPT-5.4 $2.50 cache $0.25 $15.00 $1.80 same blend 1.05M Broadly available balanced OpenAI tier
GPT-5.4 mini $0.75 cache $0.075 $4.50 $0.541 same blend 400K Lower-cost OpenAI subagents
Claude Opus 4.8 $5.00 cache $0.50 $25.00 $3.41 same blend 1M Anthropic Opus frontier alternative
Gemini 3.5 Flash $1.50 cache $0.15 $9.00 $1.08 same blend 1M Google multimodal context alternative
§ 05 / DEEP LINKS

Specific scenarios.

All calculators →

Frequently asked.

Practical GPT-5.6 Sol pricing questions, with OpenAI's published rates separated from workload assumptions.

Q · 01 How much does GPT-5.6 Sol cost? +
OpenAI lists GPT-5.6 Sol at $4/M input tokens, $0.40/M cached input tokens, $5/M cache writes and $20/M output tokens in the standard short-context tier. That is a cut from $5 / $0.50 / $6.25 / $30, and the page marks it promotional at least through November 21, 2026. Verified on OpenAI's pricing docs 2026-08-23.
Q · 02 Is GPT-5.6 Sol generally available? +
Yes. OpenAI made the GPT-5.6 series (Sol, Terra, Luna) generally available on July 9, 2026 across ChatGPT, Codex, and the API, after an initial limited preview that opened June 26, 2026.
Q · 03 What is the context window? +
OpenAI's model docs list a 1.05M-token context window and 128K max output for GPT-5.6 Sol, with a stated knowledge cutoff of February 16, 2026.
Q · 04 How much cheaper are Batch and Flex? +
OpenAI lists Batch and Flex for GPT-5.6 Sol at half the standard short-context rate: $2/M input, $0.20/M cached input, $2.50/M cache writes and $10/M output. Fast mode - the tier renamed from Priority on July 30, 2026 - is double standard: $8/M input, $0.80/M cached input, $10/M cache writes and $40/M output.
Q · 05 When does long-context pricing apply? +
For requests above 272K input tokens, GPT-5.6 Sol uses the long-context tier: $8/M input, $0.80/M cached input, $10/M cache writes and $30/M output — double the standard input rate and 1.5x the standard output rate, the same multipliers as before the cut. The calculator on this page uses the standard short-context tier.
Q · 06 Which GPT-5.6 tier should I choose? +
Use Sol for the hardest tasks, Terra for balanced agent workloads, and Luna for lower-cost routing or high-volume production. In the shared 92/8 agentic blend with 82% cache hits, Sol now lands near $2.73/M, Terra near $1.52/M, and Luna near $0.15/M. The August cut narrowed the family: Luna is 17.9x cheaper than Sol, down from 25x, and Sol is only 1.8x Terra rather than 2.5x - which is exactly the sort of change that makes an old routing rule wrong.
Q · 07 Did Sol get cheaper in the 2026 price cuts? +
Not in July, but yes in August. The July 30, 2026 cut took Luna down 80% and Terra 20% and left Sol at its launch rate, widening the gap to 25x. On or before August 23 Sol had its own cut - to $4 / $0.40 / $5 / $20 - which brings the blended agentic rate to $2.73/M and narrows the gap to Luna to 17.9x. Unlike July, this one comes labelled promotional, guaranteed at least through November 21, 2026, with no published rate for what follows.