Pricing · Alibaba (Qwen)

Qwen3.8-Omni-Flash

Alibaba (Qwen)'s qwen3.8-omni-flash, 1M tokens of context. Every number on this page was read off Alibaba (Qwen)'s own pricing page — never estimated, never recalled from memory.

How Alibaba (Qwen) describes it

“A model for audio and video understanding and content analysis, with text, image, audio, and video input and text output.”

From Alibaba (Qwen)'s documentation

Pricing

Input $0.15 per 1M tokens
Output $0.47 per 1M tokens
Cached input $0.016 per 1M tokens · 0.107× the input price
Region Singapore (International) the prices above; Beijing and Alibaba's Global regions (Frankfurt, Virginia, Tokyo, Hong Kong) charge $0.113 input and $0.382 output

Alibaba (Qwen) prices by region. This page shows the Singapore (International) region: the one Alibaba (Qwen) labels international and the one resellers mirror. Beijing and Alibaba's Global regions (Frankfurt, Virginia, Tokyo, Hong Kong) charge $0.113 per 1M input tokens and $0.382 per 1M output tokens (cached input $0.014). Same source page, same date.

At the calculator's starting volume, 1M input tokens and 200K output tokens with no cache hits, Qwen3.8-Omni-Flash costs $0.244: $0.15 for input and $0.094 for output, at the prices above. Same formula as the calculator; change the volume there.

Where this price sits

$0.15 per 1M input tokens is the 6th cheapest input price of the 67 current models with a verified price on this site, and the 2nd of Alibaba (Qwen)'s 6. Output costs 3.1× the input price.

Specifications

API model id qwen3.8-omni-flash
Provider Alibaba (Qwen)
Context window 1M tokens
Max output 131K tokens
Status Current

It returns up to 131K tokens per response, against a context window of 1M tokens.

Questions about Qwen3.8-Omni-Flash pricing

How much does Qwen3.8-Omni-Flash cost per million tokens?
$0.15 per 1M input tokens and $0.47 per 1M output tokens in the Singapore (International) region, read on Alibaba (Qwen)'s official pricing page on September 28, 2026.
What does a cached input token cost on Qwen3.8-Omni-Flash?
$0.016 per 1M cached input tokens, 10.7% of the $0.15 input price.

Omni-modal model with text output: the model page (alibabacloud.com/help/en/model-studio/qwen3-8-omni-flash, 'Last Updated: Sep 18, 2026') states 'text, image, audio, and video input and text output', 'Context window | 1M tokens', 'Maximum output length | 131072 tokens', and prints no price ('See Model pricing for inference prices'). Prices read by hand on 2026-09-28 from Alibaba's Model Studio pricing page, section Qwen-Omni ('Pricing rule: billed by input tokens and output tokens'), Singapore block: 'qwen3.8-omni-flash | International | USD 0.15 | USD 0.016 | USD 0.47' (input, cache-hit input, output per million tokens); Beijing and the four Global regions: 'USD 0.113 | USD 0.014 | USD 0.382'. Same figures as Qwen3.8-Flash. Token counting for audio and video input follows Alibaba's billing rules, not shown here. OpenRouter listed it on September 21, 2026; no release date is printed.

Verified September 28, 2026 against Alibaba (Qwen)'s official pricing page. Confidence: confirmed. No price change recorded for this model. How the data is checked.

Other Alibaba (Qwen) models with a verified price

Cheaper or equally priced alternatives

Current models with an input price at or below $0.15 per 1M tokens.