DeepSeek V4-Flash
Pricing as of July 2026$0.140/1M input · $0.280/1M output · 1.0M context · DeepSeek
Count Tokens
—
Tokens
—
Input Cost
—
Output Cost
Estimate Monthly Cost
Monthly Cost Estimator
Quick:
<$0.0001/mo
Pick a preset above or enter custom usage
Alternatives to DeepSeek V4-Flash
Pricing Details
DeepSeek284B total params, 13B active (MoE). Cache hit: $0.0028/1M. Three reasoning modes. Replaces deepseek-chat and deepseek-reasoner aliases (deprecated 2026-07-24). From the mid-July 2026 GA launch, peak-hour pricing doubles all rates during 09:00-12:00 and 14:00-18:00 Beijing time (peak ~$0.29/$0.56); off-peak stays at standard.
Input / 1M tokens
$0.14
Output / 1M tokens
$0.28
Context Window
1.0M
Max Output
384K
Price History
Launched at current price on 2026-04-24. No price changes recorded.
Frequently Asked Questions
Common questions about DeepSeek V4-Flash pricing and usage
Read More About DeepSeek V4-Flash
DeepSeek switched off deepseek-chat and deepseek-reasoner yesterday. The rename is one line; the reasoning default it flips on is what moves your bill.
July 25, 2026 · 8 min read
GLM-4.7-flash sits at $0.07 input on AWS Bedrock and Vertex. Most coverage skipped this one.
May 7, 2026 · 8 min read
Mistral Medium 3.5 charges 17x more than DeepSeek V4 Flash and loses the only benchmark they both report
May 6, 2026 · 9 min read