Pricing · Alibaba (Qwen)

Qwen3.7-Flash

Alibaba (Qwen)'s qwen3.7-flash, 1M tokens of context. Every number on this page was read off Alibaba (Qwen)'s own pricing page — never estimated, never recalled from memory.

Pricing

Input $0.03 per 1M tokens
Output $0.13 per 1M tokens
Cached input $0.006 per 1M tokens · 0.2× the input price
Region Singapore (International) the prices above; Beijing and Alibaba's Global regions (Frankfurt, Virginia, Tokyo, Hong Kong) charge $0.028 input and $0.11 output

Alibaba (Qwen) prices by region. This page shows the Singapore (International) region: the one Alibaba (Qwen) labels international and the one resellers mirror. Beijing and Alibaba's Global regions (Frankfurt, Virginia, Tokyo, Hong Kong) charge $0.028 per 1M input tokens and $0.11 per 1M output tokens (cached input $0.006); above 32000 tokens per prompt, $0.083 and $0.33; above 256000 tokens per prompt, $0.165 and $0.66. Same source page, same date.

Long-prompt tiers

Pricing changes above a prompt-size threshold. If your average request is near the boundary, the effective rate is not the headline rate.

Condition Input Output Cached input
prompt > 32000 tokens $0.10 $0.40 $0.02
prompt > 256000 tokens $0.20 $0.80 $0.04

At the calculator's starting volume, 1M input tokens and 200K output tokens with no cache hits, Qwen3.7-Flash costs $0.056: $0.03 for input and $0.026 for output, at the prices above. Same formula as the calculator; change the volume there.

Where this price sits

$0.03 per 1M input tokens (for prompts up to 32K tokens) is the 1st cheapest input price of the 67 current models with a verified price on this site, and the 1st of Alibaba (Qwen)'s 6. Output costs 4.3× the input price.

Mentioned in

One post on this site cites Qwen3.7-Flash. The sentence is the post's own, taken from where it links here.

Recorded price changes

Every change this site has recorded for Qwen3.7-Flash, with the date it took or takes effect and why. Corrections of our own reading are listed too, and say so.

Every price change recorded on this site →

Specifications

API model id qwen3.7-flash
Provider Alibaba (Qwen)
Context window 1M tokens
Max output 131K tokens
Status Current

It returns up to 131K tokens per response, against a context window of 1M tokens.

Questions about Qwen3.7-Flash pricing

How much does Qwen3.7-Flash cost per million tokens?
$0.03 per 1M input tokens and $0.13 per 1M output tokens in the Singapore (International) region for prompts up to 32,000 tokens, read on Alibaba (Qwen)'s documentation page on September 17, 2026.
What does a cached input token cost on Qwen3.7-Flash?
$0.006 per 1M cached input tokens, 20% of the $0.03 input price.
Does Qwen3.7-Flash cost more for long prompts?
Yes. For prompts over 32,000 tokens it costs $0.10 per 1M input tokens and $0.40 per 1M output tokens, $0.02 for cached input, against $0.03 and $0.13 below that.

Added 2026-09-17 (our omission): same generation as Qwen3.7 Plus and Max, found through the new-models scan's OpenRouter signal and read on Alibaba's own Model Studio page that day. Singapore block, 'Scope: International', three input-length bands: 'Input<=32k … Input | 0.03 … Output | 0.13 … Input(Implicit Cache) | 0.006', '32k<Input<=256k … Input | 0.1 … Output | 0.4 … Input(Implicit Cache) | 0.02', '256k<Input<=1m … Input | 0.2 … Output | 0.8 … Input(Implicit Cache) | 0.04' (USD per 1M tokens). Beijing and the four Global regions print 0.028 / 0.11 / 0.006, 0.083 / 0.33 / 0.017 and 0.165 / 0.66 / 0.033. 'Context Window | 1000000', 'Max Output Length | 131072'. The page lists the snapshot qwen3.7-flash-2026-07-15 and gives no release date. It states: 'This page only shows the original pricing for model API calls, excluding any limited-time promotions.'

Verified September 17, 2026 against Alibaba (Qwen)'s documentation page. Confidence: confirmed. How the data is checked.

Other Alibaba (Qwen) models with a verified price