GPT-5.6 Sol
Last verified August 9, 2026 · OpenAI pricing ↗$5.00/1M input · $30.00/1M output · 1.1M context · OpenAI
Count Tokens
—
Tokens
—
Input Cost
—
Output Cost
Estimate Monthly Cost
Monthly Cost Estimator
Quick:
<$0.0001/mo
Pick a preset above or enter custom usage
Alternatives to GPT-5.6 Sol
Pricing Details
OpenAIFlagship tier of the GPT-5.6 family, previewed June 26, 2026 to ~20 partner orgs (gated under a June 2 US frontier-model executive order). Prices identical to GPT-5.5 ($5/$30) and left UNCHANGED by the 2026-07-30 repricing that cut Luna 80% and Terra 20%, so the in-family spread widened from 5x to 25x against Luna. Only tier with the max and ultra reasoning modes; ultra runs parallel subagents at the same per-token rate but spends more tokens per task. Cached input $0.50 (90% off). Batch and Flex: $2.50/$15.00. Fast mode ($10.00 input / $1.00 cached / $60.00 output): a flat 2x on every line for up to 2.5x speed, versus the 2.5x premium GPT-5.5 Fast charges ($12.50/$75.00). This was a RENAME, not a replacement: OpenAI's changelog records Priority Processing being renamed Fast mode on 2026-07-30 (and Sol's target speed raised), service_tier accepts both 'priority' and 'fast', and no deprecation notice exists for the old value. Fast carries published SLAs: 99.9% uptime and 99% of requests above 80 output tok/s (Terra Fast 70, Luna Fast 100). Above 272K input tokens the whole request bills at 2x input / 1.5x output ($10.00/$45.00). CHANGED 2026-08-05: OpenAI's changelog records that long-context prompts over 272K can now run in Fast mode, so the two rules DO combine; the pricing page publishes Fast long-context at $20.00 input / $2.00 cached / $90.00 output, i.e. 4x input and 3x output the standard card, the dearest cell OpenAI sells. OpenAI publishes those figures but states no stacking rule in prose. Note the >272K standard input rate ($10.00) still equals the Fast short-context input rate while delivering none of the speed. Batch/Flex long context: $5.00/$22.50. Terminal-Bench 2.1: 88.8% base, 91.9% ultra. Sol on Cerebras at up to 750 tok/s from July. Artificial Analysis: Intelligence Index 59 (#4), 70M output tokens and $3,442.81 to run the index, 63.5 tok/s (rolling; read 2026-08-01), 133.04s to first answer token; Luna runs the same suite for $190.87, an 18.04x real gap against the 25x sticker gap. Rates confirmed on OpenAI's own pricing page (verified 2026-07-31).
Input / 1M tokens
$5
Output / 1M tokens
$30
Context Window
1.1M
Max Output
128K
Price History
No price change recorded since 2026-08-09, when this site started tracking this model; it launched on 2026-06-26.
Frequently Asked Questions
Common questions about GPT-5.6 Sol pricing and usage
Read More About GPT-5.6 Sol
On August 1 we wrote that nobody was going to bill you $20 per million input tokens. Four days later OpenAI let Fast mode accept long prompts, and now somebody will.
August 9, 2026 · 11 min read
Qwen3.8-Max prices output 60% under Kimi K3. On the same benchmark suite the bill came in 11% lower, and the $2 sticker is not on any page Alibaba owns.
August 4, 2026 · 11 min read
Kimi K3's weights went up for free on Monday and every host still matches Moonshot's $3/$15 on demand. That is what happens when the smallest machine that boots a model is eight B300s.
July 28, 2026 · 11 min read