Pricing · Google
Gemini 3.5 Flash
Google's gemini-3.5-flash, 1.0M tokens of context. Every number on this page was read off Google's own pricing
page — never estimated, never recalled from memory.
How Google describes it
“Our earlier Flash model, built for speed and foundational performance across routine, high-throughput workloads.”
Pricing
| Input | $1.50 | per 1M tokens |
|---|---|---|
| Output | $9.00 | per 1M tokens |
| Cached input | $0.15 | per 1M tokens · 0.1× the input price |
At the calculator's starting volume, 1M input tokens and 200K output tokens with no cache hits, Gemini 3.5 Flash costs $3.30: $1.50 for input and $1.80 for output, at the prices above. Same formula as the calculator; change the volume there.
Where this price sits
$1.50 per 1M input tokens is the 36th cheapest input price of the 67 current models with a verified price on this site, and the 10th of Google's 11. Output costs 6.0× the input price.
Other current models in the Gemini Flash family:
- Gemini 2.5 Flash 1.0M tokens $0.30 in · $2.50 out
- Gemini 3.8 Flash 1.0M tokens $0.75 in · $3.75 out · promo
- Gemini 3.7 Flash 1.0M tokens $0.75 in · $3.75 out · promo
- Gemini 3.6 Flash 1.0M tokens $0.75 in · $3.75 out · promo
Against Gemini 2.5 Flash
Gemini 2.5 Flash is the earlier version of the same name in this dataset (still current). Against Gemini 2.5 Flash, Gemini 3.5 Flash has input at $1.50 against $0.30, 5× as much; output at $9.00 against $2.50, 3.6× as much; cached input at $0.15 against $0.03, 5× as much (per 1M tokens). Its prices were verified October 1, 2026: Gemini 2.5 Flash pricing →
Specifications
| API model id | gemini-3.5-flash | |
|---|---|---|
| Provider | ||
| Family | Gemini Flash | |
| Context window | 1.0M tokens | |
| Max output | 66K tokens | |
| Released | May 1, 2026 | |
| Status | Current | |
Gemini 3.5 Flash was released on May 1, 2026. It returns up to 66K tokens per response, against a context window of 1.0M tokens.
Questions about Gemini 3.5 Flash pricing
- How much does Gemini 3.5 Flash cost per million tokens?
- $1.50 per 1M input tokens and $9.00 per 1M output tokens, read on Google's official pricing page on October 1, 2026.
- What does a cached input token cost on Gemini 3.5 Flash?
- $0.15 per 1M cached input tokens, 10% of the $1.50 input price.
Google documents it as its legacy Flash model for routine, high-throughput workloads; no shutdown date has been announced.
Verified October 1, 2026 against Google's official pricing page. Confidence: confirmed. No price change recorded for this model. How the data is checked.
Other Google models with a verified price
- Gemini 2.5 Flash-Lite Gemini Flash-Lite $0.10 in · $0.40 out
- Gemini 3.5 Flash-Lite Gemini Flash-Lite $0.30 in · $2.50 out
- Gemini 2.5 Flash Gemini Flash $0.30 in · $2.50 out
- Gemini 3.8 Flash Gemini Flash $0.75 in · $3.75 out · promo
- Gemini 3.7 Flash Gemini Flash $0.75 in · $3.75 out · promo
- Gemini 3.6 Flash Gemini Flash $0.75 in · $3.75 out · promo
- Gemini Robotics ER 2 Preview Gemini Robotics $1.00 in · $5.00 out · promo
- Gemini Robotics ER 2 Streaming Preview Gemini Robotics $1.00 in · $5.00 out · promo
- Gemini 2.5 Pro Gemini Pro $1.25 in · $10.00 out
- Gemini 3.1 Pro Preview Gemini Pro $2.00 in · $12.00 out
Cheaper or equally priced alternatives
Current models with an input price at or below $1.50 per 1M tokens.
- Mistral Medium 3.5 Mistral $1.50 in · $7.50 out
- GLM 5.2 Mistral $1.40 in · $4.40 out
- GLM 5.3 Mistral $1.40 in · $4.40 out
- DeepSeek V4 Pro DeepSeek $1.32 in · $3.96 out
Compare every model in one table →
Work out what this costs at your volume →