Pricing · Alibaba (Qwen)
Qwen3.8-Flash
Alibaba (Qwen)'s qwen3.8-flash, 1M tokens of context. Every number on this page was read off Alibaba (Qwen)'s own pricing
page — never estimated, never recalled from memory.
How Alibaba (Qwen) describes it
“Qwen3.8-Flash is the latest multimodal model from the Qwen family, combining powerful reasoning and generation with remarkable speed.”
Pricing
| Input | $0.15 | per 1M tokens |
|---|---|---|
| Output | $0.47 | per 1M tokens |
| Cached input | $0.016 | per 1M tokens · 0.107× the input price |
| Region | Singapore (International) | the prices above; Beijing and Alibaba's Global regions (Frankfurt, Virginia, Tokyo, Hong Kong) charge $0.113 input and $0.382 output |
Alibaba (Qwen) prices by region. This page shows the Singapore (International) region: the one Alibaba (Qwen) labels international and the one resellers mirror. Beijing and Alibaba's Global regions (Frankfurt, Virginia, Tokyo, Hong Kong) charge $0.113 per 1M input tokens and $0.382 per 1M output tokens (cached input $0.014). Same source page, same date. The reseller price this site tracks for this model matches the Singapore (International) figures rather than the Beijing and Global ones: OpenRouter's Alibaba endpoint at $0.15 in and $0.47 out, checked October 1, 2026.
At the calculator's starting volume, 1M input tokens and 200K output tokens with no cache hits, Qwen3.8-Flash costs $0.244: $0.15 for input and $0.094 for output, at the prices above. Same formula as the calculator; change the volume there.
Where this price sits
$0.15 per 1M input tokens is the 6th cheapest input price of the 67 current models with a verified price on this site, and the 2nd of Alibaba (Qwen)'s 6. Output costs 3.1× the input price.
Mentioned in
One post on this site cites Qwen3.8-Flash. The sentence is the post's own, taken from where it links here.
-
Prompt Caching Is Not 10% Everywhere: What a Cache Hit Costs Across Eight Providers · September 16, 2026
In the Singapore region this site publishes, Qwen3.8 Max is $0.25 cached on $2.00 of input (0.125×), Qwen3.8 Flash $0.016 on $0.15 (0.107×), Qwen3.7 Plus $0.08 on $0.40 and Qwen3.7 Max $0.50 on $2.50 (0.2×), all verified September 11, 2026, and Qwen3.7 Flash $0.006 on $0.03 (0.2×), added and verified September 17.
Recorded price changes
Every change this site has recorded for Qwen3.8-Flash, with the date it took or takes effect and why. Corrections of our own reading are listed too, and say so.
-
September 11, 2026 · change of this site's criterion
Change of criterion on our side, not a price change by Alibaba: no Alibaba price rose. Alibaba Cloud's Model Studio pages list the same model at two price levels: the China (Beijing) region and the Global regions (Frankfurt, Virginia, Tokyo, Hong Kong) share one level, and the Singapore region, labelled by Alibaba "Scope: International", is higher. Until September 11, 2026 the first-party table quoted the Beijing/Global level while the provider comparison quoted Alibaba's OpenRouter endpoint at the Singapore level, so the site showed two different "Alibaba" prices for the same model. From September 11 the reference region is Singapore (International), the one Alibaba labels international, the one every reseller we track (OpenRouter, Together, Fireworks) mirrors, and the higher of the two; the Beijing/Global figures stay published next to it. Published figures move from 1.65 / 4.951 to 2.00 / 6.00 (Qwen3.8 Max), 0.113 / 0.382 to 0.15 / 0.47 (Qwen3.8 Flash), 0.276 / 1.101 to 0.40 / 1.60 (Qwen3.7 Plus, prompts up to 256K; the tier above 256K, 1.20 / 4.80, is now recorded too) and 1.65 / 4.951 to 2.50 / 7.50 (Qwen3.7 Max). Alibaba's pages state that they show the original pricing excluding any limited-time promotions.
Every price change recorded on this site →
Where else this model is served
One other host serves Qwen3.8-Flash. Across Alibaba (Qwen) and that host the input price spans 1.0×, within the 1.15× that this site treats as the same price. Each figure links to the host's own pricing page, with the date it was checked.
| Host | Input | Output | Checked |
|---|---|---|---|
| Alibaba (Qwen) (direct) | $0.15 | $0.47 | September 11, 2026 |
| OpenRouter (via Alibaba) | $0.15 | $0.47 | October 1, 2026 |
How every host prices every open model →
What the OpenRouter price means, and what a request actually pays →
Against Qwen3.7-Flash
Qwen3.7-Flash is the earlier version of the same name in this dataset (still current). Against Qwen3.7-Flash, Qwen3.8-Flash has input at $0.15 against $0.03, 5× as much; output at $0.47 against $0.13, 3.6× as much; cached input at $0.016 against $0.006, 2.7× as much (per 1M tokens). Its prices were verified September 17, 2026: Qwen3.7-Flash pricing →
Specifications
| API model id | qwen3.8-flash | |
|---|---|---|
| Provider | Alibaba (Qwen) | |
| Context window | 1M tokens | |
| Max output | 131K tokens | |
| Released | August 28, 2026 | |
| Status | Current | |
Qwen3.8-Flash was released on August 28, 2026. It returns up to 131K tokens per response, against a context window of 1M tokens.
Questions about Qwen3.8-Flash pricing
- How much does Qwen3.8-Flash cost per million tokens?
- $0.15 per 1M input tokens and $0.47 per 1M output tokens in the Singapore (International) region, read on Alibaba (Qwen)'s documentation page on September 11, 2026.
- What does a cached input token cost on Qwen3.8-Flash?
- $0.016 per 1M cached input tokens, 10.7% of the $0.15 input price.
Verified September 11, 2026 against Alibaba (Qwen)'s documentation page. Confidence: confirmed. How the data is checked.
Other Alibaba (Qwen) models with a verified price
- Qwen3.7-Flash $0.03 in · $0.13 out
- Qwen3.8-Omni-Flash $0.15 in · $0.47 out
- Qwen3.7-Plus $0.40 in · $1.60 out
- Qwen3.8-Max $2.00 in · $6.00 out
- Qwen3.7-Max $2.50 in · $7.50 out
Cheaper or equally priced alternatives
Current models with an input price at or below $0.15 per 1M tokens.
- Mistral Small 4 Mistral $0.15 in · $0.60 out
- Ministral 3 8B Mistral $0.15 in · $0.15 out
- Command R (08-2024) Cohere $0.15 in · $0.60 out
- Qwen3.8-Omni-Flash Alibaba (Qwen) $0.15 in · $0.47 out
Compare every model in one table →
Work out what this costs at your volume →