GPT-5.6 Luna
Last verified August 17, 2026 · OpenAI pricing ↗$0.200/1M input · $1.20/1M output · 1.1M context · OpenAI
Count Tokens
—
Tokens
—
Input Cost
—
Output Cost
Estimate Monthly Cost
Monthly Cost Estimator
Quick:
<$0.0001/mo
Pick a preset above or enter custom usage
Alternatives to GPT-5.6 Luna
Pricing Details
OpenAICheapest, fastest tier of the GPT-5.6 family, for routine latency- and cost-sensitive traffic. Price cut 80% on 2026-07-30 from $1.00/$6.00; every line fell by exactly 5x. Cached input: $0.02/1M. Batch and Flex: $0.10/$0.60. Fast mode (Priority Processing renamed 2026-07-30; both service_tier values still work): $0.40 input / $0.04 cached / $2.40 output, a flat 2x. Carries the highest published latency floor in the family, 99% of requests above 100 output tok/s, against Sol Fast's 80 despite costing 25x less on output. Above 272K input tokens the whole request bills at 2x input / 1.5x output ($0.40/$1.80). CHANGED 2026-08-05: Fast mode now accepts long-context prompts, and the pricing page publishes Fast long context at $0.80 input / $0.08 cached / $3.60 output. Batch/Flex long context: $0.20/$0.90. Now lists below OpenAI's own GPT-5.4 Nano ($0.20/$1.25). Knowledge cutoff Feb 16, 2026. Chat Completions and Batch only; no realtime, fine-tuning or embeddings. Artificial Analysis (max reasoning effort): Intelligence Index 51, 130M output tokens and $190.87 to run the index, 184.4 tok/s, 116.50s TTFT. Rates confirmed on OpenAI's own pricing page (verified 2026-07-31).
Input / 1M tokens
$0.2
Output / 1M tokens
$1.2
Context Window
1.1M
Max Output
128K
Price History
No price change recorded since 2026-08-09, when this site started tracking this model; it launched on 2026-06-26.
Frequently Asked Questions
Common questions about GPT-5.6 Luna pricing and usage
Read More About GPT-5.6 Luna
OpenAI says the Agents API has no additional fees, and it does not. The sandbox every hosted session runs in bills at a container card whose rates have not moved since March, on a memory tier you cannot pick and a clock OpenAI has not said when it starts.
September 13, 2026 · 15 min read
OpenAI priced GPT-Live-1 at $0.05 a minute, the number Vapi charges to host a voice agent with no model in it. The meter runs through silence, hold music and every second the backend is thinking, and the backend is a second bill.
September 11, 2026 · 12 min read
DeepSeek cut Flash to $0.15 this morning and will bill its Pro tier at the same rate from Monday. For the 96 hours in between, deepseek-v4-pro costs 4.4x more than the model DeepSeek says has already surpassed it.
September 10, 2026 · 13 min read