GPT-5.4 Mini
Last verified July 24, 2026 · OpenAI pricing ↗$0.750/1M input · $4.50/1M output · 400K context · OpenAI
mid-rangehigh-volumechatbotscodingreal-time
When to use GPT-5.4 Mini
The sweet spot for high-volume applications. 70% cheaper than GPT-5.4 with surprisingly close performance on most tasks. Use when you need thousands of API calls per day without burning through budget.
- 3x cheaper than GPT-5.4 for input tokens
- 400K context window covers most use cases
- Fast response times for real-time applications
- Supports cached input pricing at $0.075/1M
- Computer use and tool calling support
Count Tokens
—
Tokens
—
Input Cost
—
Output Cost
Estimate Monthly Cost
Monthly Cost Estimator
Quick:
<$0.0001/mo
Pick a preset above or enter custom usage
Alternatives to GPT-5.4 Mini
Pricing Details
OpenAISupports computer use, tool calling. 2x faster than GPT-5 mini. Cached input: $0.075/1M. Batch: $0.375/$2.25.
Input / 1M tokens
$0.75
Output / 1M tokens
$4.5
Context Window
400K
Max Output
128K
Price History
Launched at current price. No price changes recorded.
Frequently Asked Questions
Common questions about GPT-5.4 Mini pricing and usage
Read More About GPT-5.4 Mini
Microsoft's MAI-Code-1-Flash matches GPT-5.4 Mini to the cent, then claims a third fewer tokens. You still can't call it through an API.
June 6, 2026 · 8 min read
Mistral Small 4: $0.15 per million input tokens for a multimodal MoE model
March 23, 2026 · 7 min read
GPT-5.4 Mini vs Nano: pricing, benchmarks, and which one to use
March 23, 2026 · 9 min read