Pricing · OpenAI
GPT-6 Astra
OpenAI's gpt-6-astra, 1.1M tokens of context. Every number on this page was read off OpenAI's own pricing
page — never estimated, never recalled from memory.
How OpenAI describes it
“GPT-6 Astra is our most capable model for the most demanding work.”
Pricing
| Input | $10.00 | per 1M tokens |
|---|---|---|
| Output | $50.00 | per 1M tokens |
| Cached input | $1.00 | per 1M tokens · 0.1× the input price |
| Cache write | $12.50 | per 1M tokens |
Long-prompt tiers
Pricing changes above a prompt-size threshold. If your average request is near the boundary, the effective rate is not the headline rate.
| Condition | Input | Output | Cached input |
|---|---|---|---|
| prompt > 272000 tokens | $20.00 | $75.00 | $2.00 |
At the calculator's starting volume, 1M input tokens and 200K output tokens with no cache hits, GPT-6 Astra costs $20.00: $10.00 for input and $10.00 for output, at the prices above. Same formula as the calculator; change the volume there.
Where this price sits
$10.00 per 1M input tokens (for prompts up to 272K tokens) is the 62nd cheapest input price of the 67 current models with a verified price on this site, and the 10th of OpenAI's 11. Output costs 5.0× the input price.
Other current models in the GPT-6 family:
- GPT-6 Luna 1.1M tokens $0.10 in · $0.50 out
- GPT-6.1 Sol 1.1M tokens $2.00 in · $10.00 out
- GPT-6 Sol 1.1M tokens $2.00 in · $10.00 out
Other price lists OpenAI publishes for GPT-6 Astra
The prices above are OpenAI's Standard list, the one the table and the calculator use. OpenAI also sells GPT-6 Astra on 4 other lists, each with a condition. Figures and wording are quoted from OpenAI's own pages as read on the date shown; this site computes none of them.
-
Batch · half of Standard on every column · wait: results within 24 hours
Printed for GPT-6 Astra: $5.00 input, $0.50 cached input, $6.25 cache write, $25.00 output, per 1M tokens.
Learn how to use OpenAI's Batch API to send asynchronous groups of requests with 50% lower costs, a separate pool of significantly higher rate limits, and a clear 24-hour turnaround time.
Batches that do not complete in time eventually move to an expired state; unfinished requests within that batch are cancelled
Source, read September 16, 2026
-
Flex · half of Standard on every column, the same figures as Batch · accept slower responses and occasional unavailability
Printed for GPT-6 Astra: $5.00 input, $0.50 cached input, $6.25 cache write, $25.00 output, per 1M tokens.
Flex processing provides lower costs for Responses or Chat Completions requests in exchange for slower response times and occasional resource unavailability.
Due to slower processing speeds with Flex processing, request timeouts are more likely.
Source, read September 16, 2026
-
Fast mode · double Standard on every column · pay more to be scheduled first
Printed for GPT-6 Astra: $20.00 input, $2.00 cached input, $25.00 cache write, $100.00 output, per 1M tokens.
Priority processing was renamed Fast mode on July 30, 2026.
If your traffic ramps too fast, the system may downgrade some Fast mode requests to standard speeds and charge standard rates.
Fast mode is unavailable for GPT-6 Astra with EU data residency.
Source, read September 16, 2026
-
Ultrafast · the figures printed on the pricing page's Ultrafast tab · pay more for the fastest output, GPT-6 Astra only
Printed for GPT-6 Astra: $60.00 input, $6.00 cached input, $75.00 cache write, $300.00 output, per 1M tokens.
Ultrafast mode is the fastest service tier in the OpenAI API. It is broadly available for GPT-6 Astra, with preview access for GPT-5.6 Sol. Use it when speed justifies the higher cost.
Ultrafast supports US data residency and global processing only. It does not support EU or other non-US regional processing endpoints.
Source, read September 30, 2026
Mentioned in
3 posts on this site cite GPT-6 Astra. The sentence is the post's own, taken from where it links here.
-
Same List Price, Different Bill: Five Pairs of LLM APIs Where the Workload Decides What You Pay · September 30, 2026
GPT-6 Astra and Claude Fable 5.1 both list at $10.00 / $50.00 (both verified September 28, 2026).
-
Prompt Caching Is Not 10% Everywhere: What a Cache Hit Costs Across Eight Providers · September 16, 2026
OpenAI, all eight current models: GPT-6 Astra $1.00 on $10.00, GPT-5.6 Sol $0.40 on $4.00, Terra $0.20 on $2.00, Luna $0.02 on $0.20, GPT-5.3 Codex $0.175 on $1.75, ChatGPT (chat-latest) $0.50 on $5.00, GPT-5.6 Cyber $1.25 on $12.50 and the access-restricted GPT-Rosalind $0.50 on $5.00, verified September 14, 2026, and September 16 for Astra and Rosalind.
-
The Second Price List: What Batch, Flex, Priority and Data-Sharing Discounts Cost Across Seven Providers · September 16, 2026
On GPT-6 Astra, $10.00 in and $50.00 out become $5.00 and $25.00; cached input drops from $1.00 to $0.50 and the cache write from $12.50 to $6.25.
Recorded price changes
Every change this site has recorded for GPT-6 Astra, with the date it took or takes effect and why. Corrections of our own reading are listed too, and say so.
-
September 16, 2026 · added to this dataset, our omission
Omission on our side, not a price change by any provider. On September 16, 2026, while reading OpenAI's pricing page by hand for a check on batch and flex prices, GPT-6 Astra turned up in the Standard table with no row in this dataset. A hand sweep of the first-party pricing pages of OpenAI, Anthropic, Google, xAI, Mistral, DeepSeek, Alibaba and Cohere against the dataset then found eight models charged per token with text output and a published price that had no row: GPT-6 Astra and GPT-Rosalind (OpenAI), Claude Haiku 3.5 (Anthropic, retired on the first-party API but still priced), Gemini 2.5 Computer Use Preview, Gemini Robotics ER 2 Preview and its Streaming variant, and Gemini Omni Flash Preview, a separate endpoint id from the generally available Omni Flash row at the same prices (Google), and GLM 5.3 (served by Mistral). Astra had been in OpenAI's Standard table at least since September 10, 2026, when the weekly detector first received that page's text, and this site's OpenRouter explainer already named it from the September 15 snapshot. The weekly detector could not have caught any of these: it compares the rows that exist against the page and has no step for models the page lists and the dataset does not. Each new row carries the page quote it was read from. Audio, image, video, embedding and per-minute models found in the same sweep were left out on purpose: the table covers models charged per token with text output.
Every price change recorded on this site →
Specifications
| API model id | gpt-6-astra | |
|---|---|---|
| Provider | OpenAI | |
| Family | GPT-6 | |
| Context window | 1.1M tokens | |
| Max output | 128K tokens | |
| Status | Current | |
It returns up to 128K tokens per response, against a context window of 1.1M tokens.
Questions about GPT-6 Astra pricing
- How much does GPT-6 Astra cost per million tokens?
- $10.00 per 1M input tokens and $50.00 per 1M output tokens for prompts up to 272,000 tokens, read on OpenAI's official pricing page on September 28, 2026.
- What does a cached input token cost on GPT-6 Astra?
- $1.00 per 1M cached input tokens, 10% of the $10.00 input price; writing to the cache costs $12.50 per 1M tokens.
- Does GPT-6 Astra cost more for long prompts?
- Yes. For prompts over 272,000 tokens it costs $20.00 per 1M input tokens and $75.00 per 1M output tokens, $2.00 for cached input, against $10.00 and $50.00 below that.
Added 2026-09-16 after a hand sweep of OpenAI's pricing page found it missing from this dataset (our omission, not a new price). Pricing page, Standard tab, first row: 'gpt-6-astra $10.00 $1.00 $12.50 $50.00 $20.00 $2.00 $25.00 $75.00' (short context: input, cached input, cache writes, output; then the same four for long context). Model page (developers.openai.com/api/docs/models/gpt-6-astra, read 2026-09-16): 'Our most capable model, built for the hardest end-to-end work', '1,050,000 context window 128,000 max output tokens', 'Prompts with more than 272K input tokens are priced at 2x input and cache rates and 1.5x output for the full request. Cache writes are billed at 1.25x the uncached input token rate. Batch and Flex are priced at 50% of Standard rates. Fast mode is priced at 2x the applicable rates.' The Fast mode row ($20 / $100) is not the long-context tier. The page gives a knowledge cutoff (Apr 30, 2026), not a release date.
Verified September 28, 2026 against OpenAI's official pricing page. Confidence: confirmed. How the data is checked.
Other OpenAI models with a verified price
- GPT-6 Luna GPT-6 $0.10 in · $0.50 out
- GPT-5.6 Luna GPT-5.6 $0.20 in · $1.20 out
- GPT-5.3 Codex Codex $1.75 in · $14.00 out
- GPT-6.1 Sol GPT-6 $2.00 in · $10.00 out
- GPT-6 Sol GPT-6 $2.00 in · $10.00 out
- GPT-5.6 Terra GPT-5.6 $2.00 in · $12.00 out
- GPT-5.6 Sol GPT-5.6 $4.00 in · $20.00 out · promo
- GPT-Rosalind GPT-Rosalind $5.00 in · $25.00 out
- ChatGPT (chat-latest) GPT $5.00 in · $30.00 out
- GPT-5.6 Cyber GPT-5.6 $12.50 in · $75.00 out
Cheaper or equally priced alternatives
Current models with an input price at or below $10.00 per 1M tokens.
- Claude Fable 5.1 Anthropic $10.00 in · $50.00 out
- Claude Mythos 5.1 Anthropic $10.00 in · $50.00 out
- Claude Fable 5 Anthropic $10.00 in · $50.00 out
- Claude Mythos 5 Anthropic $10.00 in · $50.00 out
Compare every model in one table →
Work out what this costs at your volume →