Gemini 3.8 Flash
Last verified September 3, 2026 · Google pricing ↗$0.750/1M input · $3.75/1M output · 1.0M context · Google
Count Tokens
—
Tokens
—
Input Cost
—
Output Cost
Estimate Monthly Cost
Monthly Cost Estimator
Quick:
<$0.0001/mo
Pick a preset above or enter custom usage
Alternatives to Gemini 3.8 Flash
Pricing Details
GoogleGA September 2, 2026, stable on day one (no -preview suffix), tagged New Stable on Google's models index. Google's third Flash release in six weeks after 3.6 Flash (July 21) and 3.7 Flash (August 13), and the third to carry the identical introductory card: $0.75/$3.75, cached input $0.075, cache storage $0.50/1M/hr, Batch and Flex $0.375/$1.875, Priority $1.35/$6.75. All 13 priced rows across the four tiers are byte-identical to 3.7 and 3.6 Flash, and every one of them doubles by exactly 2.0x on January 1, 2027 ($1.50/$7.50 standard), so no tier, cache-hit rate or prompt mix changes what the reversion does to a bill. The expiry is a calendar date shared across models rather than a per-model introductory window: 3.7 Flash gets 141 days of the discount, 3.8 Flash only 121. Google's launch footnote states it verbatim: 'Introductory price expires on December 31, 2026.' Vertex describes the same discount differently, as 'promotional pricing provided through 50% credits back on net spend', and applies its usual 10% non-global region surcharge ($0.825/$4.125 now, $1.65/$8.25 from January). Cache read is 0.1x base input, the constant Google holds on every Gemini Standard row and across the reversion (it loosens on cheaper tiers where cent-rounding pushes 3.5 Flash-Lite's Batch and Flex reads to 0.133x). No context-length tiers, flat across the full 1,048,576-token window. Thinking supported at low/medium/high; minimal returns an error. Output price includes thinking tokens. Batch, Flex and Priority all supported; computer use in preview. Inputs text, image, video, audio and PDF. Google-reported: HLE-Verified 54.9%; outperforms 3.7 Flash on Vals Finance Agent V2 and Harvey's Legal Agent Benchmark; DeepSWE v1.1 beats most larger frontier models at a fraction of the cost (no figure published). Google's own launch post warns the model 'might use more tokens' and says 3.7 Flash 'remains fully supported for efficiency-first workloads', so at an identical rate card 3.8 Flash can cost more per task. Knowledge cutoff not published. The sibling 3.8 Flash Cyber is gated to the Fairwind Program and has no published price, model ID or context window. API ID: gemini-3.8-flash. Estimated tokens (different tokenizer).
Input / 1M tokens
$0.75
Output / 1M tokens
$3.75
Context Window
1.0M
Max Output
66K
Introductory pricing
Scheduled price change
Gemini 3.8 Flash bills at $0.75 input / $3.75 output per 1M tokens through December 31, 2026. From January 1, 2027 the rate becomes $1.50 input / $7.50 output per 1M tokens, a 100% increase. Every figure on this page uses the rate in effect today.
Google’s published pricing ↗Price History
Launched at current price on 2026-09-02. No price changes recorded since.
Frequently Asked Questions
Common questions about Gemini 3.8 Flash pricing and usage