Pricing · Google
Gemini 2.5 Flash-Lite
Google's gemini-2.5-flash-lite, 1.0M tokens of context. Every number on this page was read off Google's own pricing
page — never estimated, never recalled from memory.
How Google describes it
“Gemini 2.5 Flash-Lite is best for high-volume classification, simple data extraction, and extremely low-latency applications where budget and speed are the primary constraints.”
From Google's documentation
- Access: “To ensure reliable performance for everyone, we are limiting access to the 2.5 models to users who have actively used them in the past.” (model page, read October 1, 2026)
- Lifecycle: “These models are not deprecated and will continue to be served until further notice through the API.” (model page, read October 1, 2026)
- For new projects: “For any new projects, use our latest models: 3.5 Flash-Lite or 3.8 Flash.” (model page, read October 1, 2026)
Pricing
| Input | $0.10 | per 1M tokens |
|---|---|---|
| Output | $0.40 | per 1M tokens |
| Cached input | $0.01 | per 1M tokens · 0.1× the input price |
| Audio input | $0.30 | per 1M tokens |
At the calculator's starting volume, 1M input tokens and 200K output tokens with no cache hits, Gemini 2.5 Flash-Lite costs $0.18: $0.10 for input and $0.08 for output, at the prices above. Same formula as the calculator; change the volume there.
Where this price sits
$0.10 per 1M input tokens is the 3rd cheapest input price of the 67 current models with a verified price on this site, and the 1st of Google's 11. Output costs 4.0× the input price.
Other current models in the Gemini Flash-Lite family:
- Gemini 3.5 Flash-Lite 1.0M tokens $0.30 in · $2.50 out
Mentioned in
One post on this site cites Gemini 2.5 Flash-Lite. The sentence is the post's own, taken from where it links here.
-
Prompt Caching Is Not 10% Everywhere: What a Cache Hit Costs Across Eight Providers · September 16, 2026
Google, all ten Gemini models with a cache price, from Gemini 2.5 Flash-Lite at $0.01 on $0.10 to Gemini 2.5 Pro at $0.125 on $1.25, September 14, and Gemini Robotics ER 2 Preview at $0.10 on $1.00, September 16.
Against Gemini 2.0 Flash-Lite
Gemini 2.0 Flash-Lite is the earlier version of the same name in this dataset (retired). Against Gemini 2.0 Flash-Lite, Gemini 2.5 Flash-Lite has input at $0.10 against $0.075, 1.3× as much; output at $0.40 against $0.30, 1.3× as much; cached input at $0.01 against $0.0187, 1.9× less (per 1M tokens). Its prices were verified September 3, 2026: Gemini 2.0 Flash-Lite pricing →
Specifications
| API model id | gemini-2.5-flash-lite | |
|---|---|---|
| Provider | ||
| Family | Gemini Flash-Lite | |
| Context window | 1.0M tokens | |
| Max output | 66K tokens | |
| Released | July 1, 2025 | |
| Status | Current | |
Gemini 2.5 Flash-Lite was released on July 1, 2025. It returns up to 66K tokens per response, against a context window of 1.0M tokens.
Questions about Gemini 2.5 Flash-Lite pricing
- How much does Gemini 2.5 Flash-Lite cost per million tokens?
- $0.10 per 1M input tokens and $0.40 per 1M output tokens, read on Google's official pricing page on October 1, 2026.
- What does a cached input token cost on Gemini 2.5 Flash-Lite?
- $0.01 per 1M cached input tokens, 10% of the $0.10 input price.
Verified October 1, 2026 against Google's official pricing page. Confidence: confirmed. No price change recorded for this model. How the data is checked.
Other Google models with a verified price
- Gemini 3.5 Flash-Lite Gemini Flash-Lite $0.30 in · $2.50 out
- Gemini 2.5 Flash Gemini Flash $0.30 in · $2.50 out
- Gemini 3.8 Flash Gemini Flash $0.75 in · $3.75 out · promo
- Gemini 3.7 Flash Gemini Flash $0.75 in · $3.75 out · promo
- Gemini 3.6 Flash Gemini Flash $0.75 in · $3.75 out · promo
- Gemini Robotics ER 2 Preview Gemini Robotics $1.00 in · $5.00 out · promo
- Gemini Robotics ER 2 Streaming Preview Gemini Robotics $1.00 in · $5.00 out · promo
- Gemini 2.5 Pro Gemini Pro $1.25 in · $10.00 out
- Gemini 3.5 Flash Gemini Flash $1.50 in · $9.00 out
- Gemini 3.1 Pro Preview Gemini Pro $2.00 in · $12.00 out
Cheaper or equally priced alternatives
Current models with an input price at or below $0.10 per 1M tokens. The cheapest of these, Qwen3.7-Flash, is 3.3× cheaper on input.
- GPT-6 Luna OpenAI $0.10 in · $0.50 out
- Ministral 3 3B Mistral $0.10 in · $0.10 out
- Command R7B Cohere $0.0375 in · $0.15 out
- Qwen3.7-Flash Alibaba (Qwen) $0.03 in · $0.13 out
Compare every model in one table →
Work out what this costs at your volume →