Pricing · Alibaba (Qwen)
Qwen3.7-Flash
Alibaba (Qwen)'s qwen3.7-flash, 1M tokens of context. Every number on this page was read off Alibaba (Qwen)'s own pricing
page — never estimated, never recalled from memory.
Pricing
| Input | $0.03 | per 1M tokens |
|---|---|---|
| Output | $0.13 | per 1M tokens |
| Cached input | $0.006 | per 1M tokens · 0.2× the input price |
| Region | Singapore (International) | the prices above; Beijing and Alibaba's Global regions (Frankfurt, Virginia, Tokyo, Hong Kong) charge $0.028 input and $0.11 output |
Alibaba (Qwen) prices by region. This page shows the Singapore (International) region: the one Alibaba (Qwen) labels international and the one resellers mirror. Beijing and Alibaba's Global regions (Frankfurt, Virginia, Tokyo, Hong Kong) charge $0.028 per 1M input tokens and $0.11 per 1M output tokens (cached input $0.006); above 32000 tokens per prompt, $0.083 and $0.33; above 256000 tokens per prompt, $0.165 and $0.66. Same source page, same date.
Long-prompt tiers
Pricing changes above a prompt-size threshold. If your average request is near the boundary, the effective rate is not the headline rate.
| Condition | Input | Output | Cached input |
|---|---|---|---|
| prompt > 32000 tokens | $0.10 | $0.40 | $0.02 |
| prompt > 256000 tokens | $0.20 | $0.80 | $0.04 |
At the calculator's starting volume, 1M input tokens and 200K output tokens with no cache hits, Qwen3.7-Flash costs $0.056: $0.03 for input and $0.026 for output, at the prices above. Same formula as the calculator; change the volume there.
Where this price sits
$0.03 per 1M input tokens (for prompts up to 32K tokens) is the 1st cheapest input price of the 67 current models with a verified price on this site, and the 1st of Alibaba (Qwen)'s 6. Output costs 4.3× the input price.
Mentioned in
One post on this site cites Qwen3.7-Flash. The sentence is the post's own, taken from where it links here.
-
Prompt Caching Is Not 10% Everywhere: What a Cache Hit Costs Across Eight Providers · September 16, 2026
In the Singapore region this site publishes, Qwen3.8 Max is $0.25 cached on $2.00 of input (0.125×), Qwen3.8 Flash $0.016 on $0.15 (0.107×), Qwen3.7 Plus $0.08 on $0.40 and Qwen3.7 Max $0.50 on $2.50 (0.2×), all verified September 11, 2026, and Qwen3.7 Flash $0.006 on $0.03 (0.2×), added and verified September 17.
Recorded price changes
Every change this site has recorded for Qwen3.7-Flash, with the date it took or takes effect and why. Corrections of our own reading are listed too, and say so.
-
September 17, 2026 · added to this dataset, our omission
Omission on our side, not a price change by Alibaba. Qwen3.7 Flash belongs to the same generation as the Qwen3.7 Plus and Max rows and had no row in this dataset. It turned up on September 17, 2026 through the OpenRouter signal of this site's read-only scan for missing models, and its prices were read that day on Alibaba's own Model Studio page: $0.03 input, $0.13 output and $0.006 cached per million tokens in the Singapore (International) region for prompts up to 32K tokens, $0.10 / $0.40 above 32K and $0.20 / $0.80 above 256K. For Alibaba the table covers the current Qwen generations, 3.7 and 3.8; the older generations Alibaba still sells, found the same day, were left out on purpose.
Every price change recorded on this site →
Specifications
| API model id | qwen3.7-flash | |
|---|---|---|
| Provider | Alibaba (Qwen) | |
| Context window | 1M tokens | |
| Max output | 131K tokens | |
| Status | Current | |
It returns up to 131K tokens per response, against a context window of 1M tokens.
Questions about Qwen3.7-Flash pricing
- How much does Qwen3.7-Flash cost per million tokens?
- $0.03 per 1M input tokens and $0.13 per 1M output tokens in the Singapore (International) region for prompts up to 32,000 tokens, read on Alibaba (Qwen)'s documentation page on September 17, 2026.
- What does a cached input token cost on Qwen3.7-Flash?
- $0.006 per 1M cached input tokens, 20% of the $0.03 input price.
- Does Qwen3.7-Flash cost more for long prompts?
- Yes. For prompts over 32,000 tokens it costs $0.10 per 1M input tokens and $0.40 per 1M output tokens, $0.02 for cached input, against $0.03 and $0.13 below that.
Added 2026-09-17 (our omission): same generation as Qwen3.7 Plus and Max, found through the new-models scan's OpenRouter signal and read on Alibaba's own Model Studio page that day. Singapore block, 'Scope: International', three input-length bands: 'Input<=32k … Input | 0.03 … Output | 0.13 … Input(Implicit Cache) | 0.006', '32k<Input<=256k … Input | 0.1 … Output | 0.4 … Input(Implicit Cache) | 0.02', '256k<Input<=1m … Input | 0.2 … Output | 0.8 … Input(Implicit Cache) | 0.04' (USD per 1M tokens). Beijing and the four Global regions print 0.028 / 0.11 / 0.006, 0.083 / 0.33 / 0.017 and 0.165 / 0.66 / 0.033. 'Context Window | 1000000', 'Max Output Length | 131072'. The page lists the snapshot qwen3.7-flash-2026-07-15 and gives no release date. It states: 'This page only shows the original pricing for model API calls, excluding any limited-time promotions.'
Verified September 17, 2026 against Alibaba (Qwen)'s documentation page. Confidence: confirmed. How the data is checked.
Other Alibaba (Qwen) models with a verified price
- Qwen3.8-Flash $0.15 in · $0.47 out
- Qwen3.8-Omni-Flash $0.15 in · $0.47 out
- Qwen3.7-Plus $0.40 in · $1.60 out
- Qwen3.8-Max $2.00 in · $6.00 out
- Qwen3.7-Max $2.50 in · $7.50 out
Compare every model in one table →
Work out what this costs at your volume →