Pricing · Alibaba (Qwen)

Qwen3.8-Flash

Alibaba (Qwen)'s qwen3.8-flash, 1M tokens of context. Every number on this page was read off Alibaba (Qwen)'s own pricing page — never estimated, never recalled from memory.

How Alibaba (Qwen) describes it

“Qwen3.8-Flash is the latest multimodal model from the Qwen family, combining powerful reasoning and generation with remarkable speed.”

Pricing

Input $0.15 per 1M tokens
Output $0.47 per 1M tokens
Cached input $0.016 per 1M tokens · 0.107× the input price
Region Singapore (International) the prices above; Beijing and Alibaba's Global regions (Frankfurt, Virginia, Tokyo, Hong Kong) charge $0.113 input and $0.382 output

Alibaba (Qwen) prices by region. This page shows the Singapore (International) region: the one Alibaba (Qwen) labels international and the one resellers mirror. Beijing and Alibaba's Global regions (Frankfurt, Virginia, Tokyo, Hong Kong) charge $0.113 per 1M input tokens and $0.382 per 1M output tokens (cached input $0.014). Same source page, same date. The reseller price this site tracks for this model matches the Singapore (International) figures rather than the Beijing and Global ones: OpenRouter's Alibaba endpoint at $0.15 in and $0.47 out, checked October 1, 2026.

At the calculator's starting volume, 1M input tokens and 200K output tokens with no cache hits, Qwen3.8-Flash costs $0.244: $0.15 for input and $0.094 for output, at the prices above. Same formula as the calculator; change the volume there.

Where this price sits

$0.15 per 1M input tokens is the 6th cheapest input price of the 67 current models with a verified price on this site, and the 2nd of Alibaba (Qwen)'s 6. Output costs 3.1× the input price.

Mentioned in

One post on this site cites Qwen3.8-Flash. The sentence is the post's own, taken from where it links here.

Recorded price changes

Every change this site has recorded for Qwen3.8-Flash, with the date it took or takes effect and why. Corrections of our own reading are listed too, and say so.

Every price change recorded on this site →

Where else this model is served

One other host serves Qwen3.8-Flash. Across Alibaba (Qwen) and that host the input price spans 1.0×, within the 1.15× that this site treats as the same price. Each figure links to the host's own pricing page, with the date it was checked.

Host Input Output Checked
Alibaba (Qwen) (direct) $0.15 $0.47 September 11, 2026
OpenRouter (via Alibaba) $0.15 $0.47 October 1, 2026

Against Qwen3.7-Flash

Qwen3.7-Flash is the earlier version of the same name in this dataset (still current). Against Qwen3.7-Flash, Qwen3.8-Flash has input at $0.15 against $0.03, 5× as much; output at $0.47 against $0.13, 3.6× as much; cached input at $0.016 against $0.006, 2.7× as much (per 1M tokens). Its prices were verified September 17, 2026: Qwen3.7-Flash pricing →

Specifications

API model id qwen3.8-flash
Provider Alibaba (Qwen)
Context window 1M tokens
Max output 131K tokens
Released August 28, 2026
Status Current

Qwen3.8-Flash was released on August 28, 2026. It returns up to 131K tokens per response, against a context window of 1M tokens.

Questions about Qwen3.8-Flash pricing

How much does Qwen3.8-Flash cost per million tokens?
$0.15 per 1M input tokens and $0.47 per 1M output tokens in the Singapore (International) region, read on Alibaba (Qwen)'s documentation page on September 11, 2026.
What does a cached input token cost on Qwen3.8-Flash?
$0.016 per 1M cached input tokens, 10.7% of the $0.15 input price.

Verified September 11, 2026 against Alibaba (Qwen)'s documentation page. Confidence: confirmed. How the data is checked.

Other Alibaba (Qwen) models with a verified price

Cheaper or equally priced alternatives

Current models with an input price at or below $0.15 per 1M tokens.