Skip to main content
TokenCost logoTokenCost

Gemini 3.5 Flash-Lite

Last verified July 26, 2026 · Google pricing

$0.300/1M input · $2.50/1M output · 1.0M context · Google

Count Tokens

Tokens
Input Cost
Output Cost

Estimate Monthly Cost

Monthly Cost Estimator

Quick:
<$0.0001/mo

Pick a preset above or enter custom usage

Alternatives to Gemini 3.5 Flash-Lite

Pricing Details

GoogleLaunched July 21, 2026 alongside Gemini 3.6 Flash. Price increase over 3.1 Flash-Lite ($0.25/$1.50): +20% input, +67% output. Same 1.2x/1.667x multipliers apply across all four service tiers. Google raised the floor because the new Lite jumps 31 to 54 on Terminal-Bench 2.1 and nearly doubles GDPval-AA v2 (642 to 1140). Cached input: $0.03/1M, cache storage $1.00/1M/hour. Batch and Flex: $0.15/$1.25. Priority: $0.54/$4.50. Audio input folded into the flat $0.30 rate, down 40% from 3.1 Flash-Lite's separate $0.50 audio line. Vertex non-global endpoints bill 10% higher ($0.33/$2.75) since July 1, 2026. Thinking on by default at minimal level; thinking tokens bill as output. Stable, no -preview suffix. In Gemini API, AI Studio, and Google Search. Knowledge cutoff March 2026. Estimated tokens (different tokenizer).
Input / 1M tokens
$0.3
Output / 1M tokens
$2.5
Context Window
1.0M
Max Output
66K

Price History

No price change recorded since 2026-07-26, when this site started tracking this model; it launched on 2026-07-21.

Frequently Asked Questions

Common questions about Gemini 3.5 Flash-Lite pricing and usage

Read More About Gemini 3.5 Flash-Lite