LLM API pricing / Google
What the Gemini API costs, model by model
Google sells 11 current models with a verified price on its API, from $0.10 to $2.00 per 1M input tokens and from $0.40 to $12.00 per 1M output tokens. Every price below was read on a page Google publishes, verified on October 1, 2026, and links to it.
Current models
| Model | Input $/1M | Output $/1M | Cached input | Context | Verified |
|---|---|---|---|---|---|
| Gemini 2.5 Flash-Lite | $0.10 | $0.40 | $0.01 | 1.0M | 2026-10-01 |
| Gemini 2.5 Flash | $0.30 | $2.50 | $0.03 | 1.0M | 2026-10-01 |
| Gemini 3.5 Flash-Lite | $0.30 | $2.50 | $0.03 | 1.0M | 2026-10-01 |
| Gemini 3.6 Flash promo | $0.75 | $3.75 | $0.075 | 1.0M | 2026-10-01 |
| Gemini 3.7 Flash promo | $0.75 | $3.75 | $0.075 | 1.0M | 2026-10-01 |
| Gemini 3.8 Flash promo | $0.75 | $3.75 | $0.075 | 1.0M | 2026-10-01 |
| Gemini Robotics ER 2 Preview promo | $1.00 | $5.00 | $0.10 | 131K | 2026-10-01 |
| Gemini Robotics ER 2 Streaming Preview promo | $1.00 | $5.00 | — | 131K | 2026-10-01 |
| Gemini 2.5 Pro | $1.25 | $10.00 | $0.125 | 1.0M | 2026-10-01 |
| Gemini 3.5 Flash | $1.50 | $9.00 | $0.15 | 1.0M | 2026-10-01 |
| Gemini 3.1 Pro Preview | $2.00 | $12.00 | $0.20 | 1.0M | 2026-10-01 |
How Google prices its models
- 10 of the 11 models publish a cached-input price, at 10% of the input price.
- 2 models charge more for long prompts (prompts over 200,000 tokens): Gemini 2.5 Pro $2.50 / $15.00, Gemini 3.1 Pro Preview $4.00 / $18.00.
- 5 prices are promotional: Gemini 3.6 Flash until December 31, 2026, then $1.50 / $7.50; Gemini 3.7 Flash until December 31, 2026, then $1.50 / $7.50; Gemini 3.8 Flash until December 31, 2026, then $1.50 / $7.50; Gemini Robotics ER 2 Preview until December 31, 2026, then $2.00 / $10.00; Gemini Robotics ER 2 Streaming Preview until December 31, 2026, then $2.00 / $10.00.
Second price lists
Prices Google prints for other ways of buying the same tokens, per 1M tokens, only where its page prints a figure for the model.
Batch · 50% of the standard interactive API cost
- Gemini 3.6 Flash: $0.375 input, $1.875 output, $0.75 input after the promotion ends, $3.75 output after the promotion ends (read September 16, 2026)
- Gemini 2.5 Pro: $0.625 input, $0.125 cached input, $5.00 output, $1.25 input above the long-prompt threshold, $7.50 output above the long-prompt threshold (read September 16, 2026)
Google: "Batch API usage is priced at 50% of the standard interactive API cost for the equivalent model." source
Flex · 50% cost reduction compared to standard rates
- Gemini 3.6 Flash: $0.375 input, $1.875 output, $0.75 input after the promotion ends, $3.75 output after the promotion ends (read September 16, 2026)
- Gemini 2.5 Pro: $0.625 input, $5.00 output, $1.25 input above the long-prompt threshold, $7.50 output above the long-prompt threshold (read September 16, 2026)
Google: "The Gemini Flex API is an inference tier that offers a 50% cost reduction compared to standard rates, in exchange for variable latency and best-effort availability." source
Priority · no multiplier printed; the page prints the prices
- Gemini 3.6 Flash: $1.35 input, $6.75 output, $2.70 input after the promotion ends, $13.50 output after the promotion ends (read September 16, 2026)
- Gemini 2.5 Pro: $2.25 input, $0.225 cached input, $18.00 output, $4.50 input above the long-prompt threshold, $27.00 output above the long-prompt threshold (read September 16, 2026)
Google: "The Gemini Priority API is a premium inference tier designed for business-critical workloads that require lower latency and the highest reliability at a premium price point." source
Retirements
Deprecated, still callable: Gemini 3.1 Flash-Lite (shutdown May 7, 2027); Gemini 3 Flash Preview (no shutdown date announced).
Retired: Gemini 2.5 Computer Use Preview (July 28, 2026), Gemini 2.0 Flash (June 1, 2026), Gemini 2.0 Flash-Lite (June 1, 2026), Gemini 3 Pro Preview (March 9, 2026).
- June 1, 2026: Shutdown of Gemini 2.0 Flash and 2.0 Flash-Lite source
- July 28, 2026: Shutdown of Gemini 2.5 Computer Use Preview (Google recommends Gemini 3.8 Flash or other Gemini 3 models with built-in computer use) source
- May 7, 2027 (announced): Shutdown of Gemini 3.1 Flash-Lite source
Recorded price changes
- January 1, 2027 · price up 2.0×. Gemini 3.6, 3.7 and 3.8 Flash pricing is promotional until December 31, 2026. On January 1, 2027 it doubles: input from $0.75 to $1.50 and output from $3.75 to $7.50 per million tokens, the list prices Google shows next to the promotional ones. source
- September 30, 2026 · change of this site's criterion · reference price withdrawn. Change of this site's criterion, not a price change by Google. Gemini Omni Flash (gemini-omni-1.1-flash) and its preview endpoint (gemini-omni-flash-preview) entered this dataset in the hand sweep of September 16, 2026, priced $1.50 per million input tokens and $9.00 per million output tokens. On September 30, 2026 Google's model page for both versions was read: it describes a model 'for fast, conversational video generation and editing' and gives its supported output as 'Video' only. The pricing table covers models charged per token with text output, so both rows left the table, the calculator and the counts that day. The $9.00 (text) figure on Google's pricing page is labelled as including thinking tokens; it is not the price of a text answer comparable with the other rows. Both rows stay in the dataset with their prices and dates, and their pages remain online, out of the search index. source
- September 16, 2026 · added to this dataset, our omission: Gemini 2.5 Computer Use Preview, Gemini Robotics ER 2 Preview, Gemini Robotics ER 2 Streaming Preview, Gemini Omni Flash Preview. Full explanation
Outside the table: Gemini Omni Flash, Gemini Omni Flash Preview, priced per token but not with text output (2026-09-30).
Analysis that cites Google models
- Prompt Caching Is Not 10% Everywhere: What a Cache Hit Costs Across Eight Providers · names 7 Google models
- Five Things Decide What You Pay for a Model, and Only One of Them Is the Model · names 4 Google models
- The Second Price List: What Batch, Flex, Priority and Data-Sharing Discounts Cost Across Seven Providers · names 2 Google models
- Same List Price, Different Bill: Five Pairs of LLM APIs Where the Workload Decides What You Pay · names 1 Google model
How every figure is checked: methodology. All providers: the full pricing table.