Skip to main content
TokenCost logoTokenCost

GPT-5.6 Luna

Last verified August 9, 2026 · OpenAI pricing

$0.200/1M input · $1.20/1M output · 1.1M context · OpenAI

Count Tokens

Tokens
Input Cost
Output Cost

Estimate Monthly Cost

Monthly Cost Estimator

Quick:
<$0.0001/mo

Pick a preset above or enter custom usage

Alternatives to GPT-5.6 Luna

Pricing Details

OpenAICheapest, fastest tier of the GPT-5.6 family, for routine latency- and cost-sensitive traffic. Price cut 80% on 2026-07-30 from $1.00/$6.00; every line fell by exactly 5x. Cached input: $0.02/1M. Batch and Flex: $0.10/$0.60. Fast mode (Priority Processing renamed 2026-07-30; both service_tier values still work): $0.40 input / $0.04 cached / $2.40 output, a flat 2x. Carries the highest published latency floor in the family, 99% of requests above 100 output tok/s, against Sol Fast's 80 despite costing 25x less on output. Above 272K input tokens the whole request bills at 2x input / 1.5x output ($0.40/$1.80). CHANGED 2026-08-05: Fast mode now accepts long-context prompts, and the pricing page publishes Fast long context at $0.80 input / $0.08 cached / $3.60 output. Batch/Flex long context: $0.20/$0.90. Now lists below OpenAI's own GPT-5.4 Nano ($0.20/$1.25). Knowledge cutoff Feb 16, 2026. Chat Completions and Batch only; no realtime, fine-tuning or embeddings. Artificial Analysis (max reasoning effort): Intelligence Index 51, 130M output tokens and $190.87 to run the index, 184.4 tok/s, 116.50s TTFT. Rates confirmed on OpenAI's own pricing page (verified 2026-07-31).
Input / 1M tokens
$0.2
Output / 1M tokens
$1.2
Context Window
1.1M
Max Output
128K

Price History

No price change recorded since 2026-08-09, when this site started tracking this model; it launched on 2026-06-26.

Frequently Asked Questions

Common questions about GPT-5.6 Luna pricing and usage

Read More About GPT-5.6 Luna