Pricing · OpenAI
GPT-6 Luna
OpenAI's gpt-6-luna, 1.1M tokens of context. Every number on this page was read off OpenAI's own pricing
page — never estimated, never recalled from memory.
How OpenAI describes it
“GPT-6 Luna is our most efficient model for focused, high-volume tasks.”
Pricing
| Input | $0.10 | per 1M tokens |
|---|---|---|
| Output | $0.50 | per 1M tokens |
| Cached input | $0.01 | per 1M tokens · 0.1× the input price |
| Cache write | $0.125 | per 1M tokens |
Long-prompt tiers
Pricing changes above a prompt-size threshold. If your average request is near the boundary, the effective rate is not the headline rate.
| Condition | Input | Output | Cached input |
|---|---|---|---|
| prompt > 272000 tokens | $0.20 | $0.75 | $0.02 |
At the calculator's starting volume, 1M input tokens and 200K output tokens with no cache hits, GPT-6 Luna costs $0.20: $0.10 for input and $0.10 for output, at the prices above. Same formula as the calculator; change the volume there.
Where this price sits
$0.10 per 1M input tokens (for prompts up to 272K tokens) is the 3rd cheapest input price of the 67 current models with a verified price on this site, and the 1st of OpenAI's 11. Output costs 5.0× the input price.
Other current models in the GPT-6 family:
- GPT-6.1 Sol 1.1M tokens $2.00 in · $10.00 out
- GPT-6 Sol 1.1M tokens $2.00 in · $10.00 out
- GPT-6 Astra 1.1M tokens $10.00 in · $50.00 out
Other price lists OpenAI publishes for GPT-6 Luna
The prices above are OpenAI's Standard list, the one the table and the calculator use. OpenAI also sells GPT-6 Luna on 3 other lists, each with a condition. Figures and wording are quoted from OpenAI's own pages as read on the date shown; this site computes none of them.
-
Batch · half of Standard on every column · wait: results within 24 hours
Printed for GPT-6 Luna: $0.05 input, $0.005 cached input, $0.0625 cache write, $0.25 output, per 1M tokens.
Learn how to use OpenAI's Batch API to send asynchronous groups of requests with 50% lower costs, a separate pool of significantly higher rate limits, and a clear 24-hour turnaround time.
Batches that do not complete in time eventually move to an expired state; unfinished requests within that batch are cancelled
Source, read September 24, 2026
-
Flex · half of Standard on every column, the same figures as Batch · accept slower responses and occasional unavailability
Printed for GPT-6 Luna: $0.05 input, $0.005 cached input, $0.0625 cache write, $0.25 output, per 1M tokens.
Flex processing provides lower costs for Responses or Chat Completions requests in exchange for slower response times and occasional resource unavailability.
Due to slower processing speeds with Flex processing, request timeouts are more likely.
Source, read September 24, 2026
-
Fast mode · double Standard on every column · pay more to be scheduled first
Printed for GPT-6 Luna: $0.20 input, $0.02 cached input, $0.25 cache write, $1.00 output, per 1M tokens.
Priority processing was renamed Fast mode on July 30, 2026.
If your traffic ramps too fast, the system may downgrade some Fast mode requests to standard speeds and charge standard rates.
Source, read September 24, 2026
Against GPT-5.6 Luna
GPT-5.6 Luna is the earlier version of the same name in this dataset (still current). Against GPT-5.6 Luna, GPT-6 Luna has input at $0.10 against $0.20, 2× less; output at $0.50 against $1.20, 2.4× less; cached input at $0.01 against $0.02, 2× less (per 1M tokens). Its prices were verified September 28, 2026: GPT-5.6 Luna pricing →
Specifications
| API model id | gpt-6-luna | |
|---|---|---|
| Provider | OpenAI | |
| Family | GPT-6 | |
| Context window | 1.1M tokens | |
| Max output | 128K tokens | |
| Knowledge cutoff | 2026-05-18 | |
| Released | September 22, 2026 | |
| Status | Current | |
GPT-6 Luna was released on September 22, 2026. It returns up to 128K tokens per response, against a context window of 1.1M tokens.
Questions about GPT-6 Luna pricing
- How much does GPT-6 Luna cost per million tokens?
- $0.10 per 1M input tokens and $0.50 per 1M output tokens for prompts up to 272,000 tokens, read on OpenAI's official pricing page on September 28, 2026.
- What does a cached input token cost on GPT-6 Luna?
- $0.01 per 1M cached input tokens, 10% of the $0.10 input price; writing to the cache costs $0.125 per 1M tokens.
- Does GPT-6 Luna cost more for long prompts?
- Yes. For prompts over 272,000 tokens it costs $0.20 per 1M input tokens and $0.75 per 1M output tokens, $0.02 for cached input, against $0.10 and $0.50 below that.
Released September 22, 2026 (OpenAI changelog, Sep 22: '… GPT-6 Luna: $0.10 input, $0.01 cached input, and $0.50 output.'). Read by hand on 2026-09-24 from OpenAI's pricing page, Flagship models, Standard tab: 'gpt-6-luna | $0.10 | $0.01 | $0.125 | $0.50 | $0.20 | $0.02 | $0.25 | $0.75' (short context: input, cached input, cache writes, output; then the same four for long context). Model page, read the same day: '1,050,000 context window', '128,000 max output tokens', 'May 18, 2026 knowledge cutoff', 'Prompts with more than 272K input tokens are priced at 2x input and cache rates and 1.5x output for the full request.' The Fast mode row ($0.20 / $1.00) is not the long-context tier.
Verified September 28, 2026 against OpenAI's official pricing page. Confidence: confirmed. No price change recorded for this model. How the data is checked.
Other OpenAI models with a verified price
- GPT-5.6 Luna GPT-5.6 $0.20 in · $1.20 out
- GPT-5.3 Codex Codex $1.75 in · $14.00 out
- GPT-6.1 Sol GPT-6 $2.00 in · $10.00 out
- GPT-6 Sol GPT-6 $2.00 in · $10.00 out
- GPT-5.6 Terra GPT-5.6 $2.00 in · $12.00 out
- GPT-5.6 Sol GPT-5.6 $4.00 in · $20.00 out · promo
- GPT-Rosalind GPT-Rosalind $5.00 in · $25.00 out
- ChatGPT (chat-latest) GPT $5.00 in · $30.00 out
- GPT-6 Astra GPT-6 $10.00 in · $50.00 out
- GPT-5.6 Cyber GPT-5.6 $12.50 in · $75.00 out
Cheaper or equally priced alternatives
Current models with an input price at or below $0.10 per 1M tokens. The cheapest of these, Qwen3.7-Flash, is 3.3× cheaper on input.
- Gemini 2.5 Flash-Lite Google $0.10 in · $0.40 out
- Ministral 3 3B Mistral $0.10 in · $0.10 out
- Command R7B Cohere $0.0375 in · $0.15 out
- Qwen3.7-Flash Alibaba (Qwen) $0.03 in · $0.13 out
Compare every model in one table →
Work out what this costs at your volume →