Pricing · Alibaba (Qwen)
Qwen3.8-Max
Alibaba (Qwen)'s qwen3.8-max, 1M tokens of context. Every number on this page was read off Alibaba (Qwen)'s own pricing
page — never estimated, never recalled from memory.
Pricing
| Input | $2.00 | per 1M tokens |
|---|---|---|
| Output | $6.00 | per 1M tokens |
| Cached input | $0.25 | per 1M tokens · 0.125× the input price |
| Region | Singapore (International) | the prices above; Beijing and Alibaba's Global regions (Frankfurt, Virginia, Tokyo, Hong Kong) charge $1.65 input and $4.951 output |
Alibaba (Qwen) prices by region. This page shows the Singapore (International) region: the one Alibaba (Qwen) labels international and the one resellers mirror. Beijing and Alibaba's Global regions (Frankfurt, Virginia, Tokyo, Hong Kong) charge $1.65 per 1M input tokens and $4.951 per 1M output tokens (cached input $0.206). Same source page, same date. The reseller prices this site tracks for this model match the Singapore (International) figures rather than the Beijing and Global ones: Together AI at $2.00 in and $6.00 out and Fireworks AI at the same, both checked October 1, 2026.
At the calculator's starting volume, 1M input tokens and 200K output tokens with no cache hits, Qwen3.8-Max costs $3.20: $2.00 for input and $1.20 for output, at the prices above. Same formula as the calculator; change the volume there.
Where this price sits
$2.00 per 1M input tokens is the 39th cheapest input price of the 67 current models with a verified price on this site, and the 5th of Alibaba (Qwen)'s 6. Output costs 3.0× the input price.
Mentioned in
3 posts on this site cite Qwen3.8-Max. The sentence is the post's own, taken from where it links here.
-
Same List Price, Different Bill: Five Pairs of LLM APIs Where the Workload Decides What You Pay · September 30, 2026
Grok 4.7 (verified September 28, 2026) and Qwen3.8-Max (verified September 11, 2026) both charge $2.00 / $6.00.
-
Prompt Caching Is Not 10% Everywhere: What a Cache Hit Costs Across Eight Providers · September 16, 2026
In the Singapore region this site publishes, Qwen3.8 Max is $0.25 cached on $2.00 of input (0.125×), Qwen3.8 Flash $0.016 on $0.15 (0.107×), Qwen3.7 Plus $0.08 on $0.40 and Qwen3.7 Max $0.50 on $2.50 (0.2×), all verified September 11, 2026, and Qwen3.7 Flash $0.006 on $0.03 (0.2×), added and verified September 17.
-
Five Things Decide What You Pay for a Model, and Only One of Them Is the Model · September 15, 2026
Qwen3.8 Max is $2.00 in and $6.00 out in Singapore against $1.65 and $4.951 elsewhere, 1.21× on input.
Recorded price changes
Every change this site has recorded for Qwen3.8-Max, with the date it took or takes effect and why. Corrections of our own reading are listed too, and say so.
-
September 11, 2026 · change of this site's criterion
Change of criterion on our side, not a price change by Alibaba: no Alibaba price rose. Alibaba Cloud's Model Studio pages list the same model at two price levels: the China (Beijing) region and the Global regions (Frankfurt, Virginia, Tokyo, Hong Kong) share one level, and the Singapore region, labelled by Alibaba "Scope: International", is higher. Until September 11, 2026 the first-party table quoted the Beijing/Global level while the provider comparison quoted Alibaba's OpenRouter endpoint at the Singapore level, so the site showed two different "Alibaba" prices for the same model. From September 11 the reference region is Singapore (International), the one Alibaba labels international, the one every reseller we track (OpenRouter, Together, Fireworks) mirrors, and the higher of the two; the Beijing/Global figures stay published next to it. Published figures move from 1.65 / 4.951 to 2.00 / 6.00 (Qwen3.8 Max), 0.113 / 0.382 to 0.15 / 0.47 (Qwen3.8 Flash), 0.276 / 1.101 to 0.40 / 1.60 (Qwen3.7 Plus, prompts up to 256K; the tier above 256K, 1.20 / 4.80, is now recorded too) and 1.65 / 4.951 to 2.50 / 7.50 (Qwen3.7 Max). Alibaba's pages state that they show the original pricing excluding any limited-time promotions.
Every price change recorded on this site →
Where else this model is served
2 other hosts serve Qwen3.8-Max. Across Alibaba (Qwen) and those hosts the input price spans 1.0×, within the 1.15× that this site treats as the same price. Each figure links to the host's own pricing page, with the date it was checked.
| Host | Input | Output | Checked |
|---|---|---|---|
| Alibaba (Qwen) (direct) | $2.00 | $6.00 | September 11, 2026 |
| Together AI | $2.00 | $6.00 | October 1, 2026 |
| Fireworks AI | $2.00 | $6.00 | October 1, 2026 |
-
September 1, 2026 · Fireworks AI (all models on this host) · price up 1.5×
Since September 1, 2026 Fireworks adds a 50% surcharge to the base price of the Serverless models it marks as US-only. The Fireworks prices on this site are the base prices, without the surcharge.
How every host prices every open model →
When comparing hosts is worth it, and when it is not →
Against Qwen3.7-Max
Qwen3.7-Max is the earlier version of the same name in this dataset (still current). Against Qwen3.7-Max, Qwen3.8-Max has input at $2.00 against $2.50, 1.3× less; output at $6.00 against $7.50, 1.3× less; cached input at $0.25 against $0.50, 2× less (per 1M tokens). Its prices were verified September 11, 2026: Qwen3.7-Max pricing →
Specifications
| API model id | qwen3.8-max | |
|---|---|---|
| Provider | Alibaba (Qwen) | |
| Context window | 1M tokens | |
| Max output | 131K tokens | |
| Released | September 2, 2026 | |
| Status | Current | |
Qwen3.8-Max was released on September 2, 2026. It returns up to 131K tokens per response, against a context window of 1M tokens.
Questions about Qwen3.8-Max pricing
- How much does Qwen3.8-Max cost per million tokens?
- $2.00 per 1M input tokens and $6.00 per 1M output tokens in the Singapore (International) region, read on Alibaba (Qwen)'s documentation page on September 11, 2026.
- What does a cached input token cost on Qwen3.8-Max?
- $0.25 per 1M cached input tokens, 12.5% of the $2.00 input price.
Verified September 11, 2026 against Alibaba (Qwen)'s documentation page. Confidence: confirmed. How the data is checked.
Other Alibaba (Qwen) models with a verified price
- Qwen3.7-Flash $0.03 in · $0.13 out
- Qwen3.8-Flash $0.15 in · $0.47 out
- Qwen3.8-Omni-Flash $0.15 in · $0.47 out
- Qwen3.7-Plus $0.40 in · $1.60 out
- Qwen3.7-Max $2.50 in · $7.50 out
Cheaper or equally priced alternatives
Current models with an input price at or below $2.00 per 1M tokens.
- GPT-6.1 Sol OpenAI $2.00 in · $10.00 out
- GPT-6 Sol OpenAI $2.00 in · $10.00 out
- GPT-5.6 Terra OpenAI $2.00 in · $12.00 out
- Claude Sonnet 5.5 Anthropic $2.00 in · $10.00 out
Compare every model in one table →
Work out what this costs at your volume →