Pricing · Google
Gemini 3.1 Flash-Lite
Google's gemini-3.1-flash-lite, 1.0M tokens of context. Every number on this page was read off Google's own pricing
page — never estimated, never recalled from memory.
How Google describes it
“Gemini 3.1 Flash-Lite is a low-latency, cost-effective multimodal model optimized for high-frequency, lightweight tasks.”
From Google's documentation
- What it is for: “The model supports text, image, video, audio, and PDF inputs, and is designed for high-volume agentic workflows, simple data extraction, and applications where latency and API cost are the primary constraints.” (model page, read October 1, 2026)
- Ways to buy it: “Batch API Supported Flex inference Supported Priority inference Supported” (model page, read October 1, 2026)
Pricing
| Input | $0.25 | per 1M tokens |
|---|---|---|
| Output | $1.50 | per 1M tokens |
| Cached input | $0.025 | per 1M tokens · 0.1× the input price |
| Audio input | $0.50 | per 1M tokens |
At the calculator's starting volume, 1M input tokens and 200K output tokens with no cache hits, Gemini 3.1 Flash-Lite costs $0.55: $0.25 for input and $0.30 for output, at the prices above. Same formula as the calculator; change the volume there.
Where this price sits
Gemini 3.1 Flash-Lite is deprecated and is not counted among the current models. Placed against the 67 current models with a verified price on this site, its $0.25 per 1M input tokens would be the 13th cheapest of 68, and the 2nd of 12 from Google. Output costs 6.0× the input price.
Other current models in the Gemini Flash-Lite family:
- Gemini 2.5 Flash-Lite 1.0M tokens $0.10 in · $0.40 out
- Gemini 3.5 Flash-Lite 1.0M tokens $0.30 in · $2.50 out
Succession
Replaced by Gemini 3.5 Flash-Lite (shutdown May 7, 2027).
Google names Gemini 3.1 Flash-Lite as the replacement for:
- Gemini 2.0 Flash-Lite retired shut down June 1, 2026
Against Gemini 2.5 Flash-Lite
Gemini 2.5 Flash-Lite is the earlier version of the same name in this dataset (still current). Against Gemini 2.5 Flash-Lite, Gemini 3.1 Flash-Lite has input at $0.25 against $0.10, 2.5× as much; output at $1.50 against $0.40, 3.8× as much; cached input at $0.025 against $0.01, 2.5× as much (per 1M tokens). Its prices were verified October 1, 2026: Gemini 2.5 Flash-Lite pricing →
Specifications
| API model id | gemini-3.1-flash-lite | |
|---|---|---|
| Provider | ||
| Family | Gemini Flash-Lite | |
| Context window | 1.0M tokens | |
| Max output | 66K tokens | |
| Released | May 1, 2026 | |
| Status | Deprecated | |
| Retires on | May 7, 2027 | |
Gemini 3.1 Flash-Lite was released on May 1, 2026. It returns up to 66K tokens per response, against a context window of 1.0M tokens.
Shutdowns at Google on record
3 dated shutdowns this site has recorded for Google, each with its source. Dates after today are announced, not past.
| Date | What | Source |
|---|---|---|
| June 1, 2026 | Shutdown of Gemini 2.0 Flash and 2.0 Flash-Lite | Source |
| July 28, 2026 | Shutdown of Gemini 2.5 Computer Use Preview (Google recommends Gemini 3.8 Flash or other Gemini 3 models with built-in computer use) | Source |
| May 7, 2027 (announced) | Shutdown of Gemini 3.1 Flash-Lite | Source |
Questions about Gemini 3.1 Flash-Lite pricing
- How much does Gemini 3.1 Flash-Lite cost per million tokens?
- $0.25 per 1M input tokens and $1.50 per 1M output tokens, read on Google's official pricing page on October 1, 2026.
- What does a cached input token cost on Gemini 3.1 Flash-Lite?
- $0.025 per 1M cached input tokens, 10% of the $0.25 input price.
- When will Gemini 3.1 Flash-Lite be retired?
- Google has set its retirement for May 7, 2027.
Verified October 1, 2026 against Google's official pricing page. Confidence: confirmed. No price change recorded for this model. How the data is checked.
Other Google models with a verified price
- Gemini 2.5 Flash-Lite Gemini Flash-Lite $0.10 in · $0.40 out
- Gemini 3.5 Flash-Lite Gemini Flash-Lite $0.30 in · $2.50 out
- Gemini 2.5 Flash Gemini Flash $0.30 in · $2.50 out
- Gemini 3.8 Flash Gemini Flash $0.75 in · $3.75 out · promo
- Gemini 3.7 Flash Gemini Flash $0.75 in · $3.75 out · promo
- Gemini 3.6 Flash Gemini Flash $0.75 in · $3.75 out · promo
- Gemini Robotics ER 2 Preview Gemini Robotics $1.00 in · $5.00 out · promo
- Gemini Robotics ER 2 Streaming Preview Gemini Robotics $1.00 in · $5.00 out · promo
- Gemini 2.5 Pro Gemini Pro $1.25 in · $10.00 out
- Gemini 3.5 Flash Gemini Flash $1.50 in · $9.00 out
- Gemini 3.1 Pro Preview Gemini Pro $2.00 in · $12.00 out
Cheaper or equally priced alternatives
Current models with an input price at or below $0.25 per 1M tokens. The cheapest of these, Ministral 3 8B, is 1.7× cheaper on input.
- GPT-5.6 Luna OpenAI $0.20 in · $1.20 out
- Ministral 3 14B Mistral $0.20 in · $0.20 out
- Mistral Small 4 Mistral $0.15 in · $0.60 out
- Ministral 3 8B Mistral $0.15 in · $0.15 out
Compare every model in one table →
Work out what this costs at your volume →