Pricing · Alibaba (Qwen)
Qwen3.7-Plus
Alibaba (Qwen)'s qwen3.7-plus, 1M tokens of context. Every number on this page was read off Alibaba (Qwen)'s own pricing
page — never estimated, never recalled from memory.
Pricing
| Input | $0.40 | per 1M tokens |
|---|---|---|
| Output | $1.60 | per 1M tokens |
| Cached input | $0.08 | per 1M tokens · 0.2× the input price |
| Region | Singapore (International) | the prices above; Beijing and Alibaba's Global regions (Frankfurt, Virginia, Tokyo, Hong Kong) charge $0.276 input and $1.101 output |
Alibaba (Qwen) prices by region. This page shows the Singapore (International) region: the one Alibaba (Qwen) labels international and the one resellers mirror. Beijing and Alibaba's Global regions (Frankfurt, Virginia, Tokyo, Hong Kong) charge $0.276 per 1M input tokens and $1.101 per 1M output tokens (cached input $0.056); above 256000 tokens per prompt, $0.826 and $3.301. Same source page, same date.
Long-prompt tiers
Pricing changes above a prompt-size threshold. If your average request is near the boundary, the effective rate is not the headline rate.
| Condition | Input | Output | Cached input |
|---|---|---|---|
| prompt > 256000 tokens | $1.20 | $4.80 | $0.24 |
At the calculator's starting volume, 1M input tokens and 200K output tokens with no cache hits, Qwen3.7-Plus costs $0.72: $0.40 for input and $0.32 for output, at the prices above. Same formula as the calculator; change the volume there.
Where this price sits
$0.40 per 1M input tokens (for prompts up to 256K tokens) is the 17th cheapest input price of the 67 current models with a verified price on this site, and the 4th of Alibaba (Qwen)'s 6. Output costs 4.0× the input price.
Mentioned in
2 posts on this site cite Qwen3.7-Plus. The sentence is the post's own, taken from where it links here.
-
Prompt Caching Is Not 10% Everywhere: What a Cache Hit Costs Across Eight Providers · September 16, 2026
In the Singapore region this site publishes, Qwen3.8 Max is $0.25 cached on $2.00 of input (0.125×), Qwen3.8 Flash $0.016 on $0.15 (0.107×), Qwen3.7 Plus $0.08 on $0.40 and Qwen3.7 Max $0.50 on $2.50 (0.2×), all verified September 11, 2026, and Qwen3.7 Flash $0.006 on $0.03 (0.2×), added and verified September 17.
-
Five Things Decide What You Pay for a Model, and Only One of Them Is the Model · September 15, 2026 · 2 mentions
Qwen3.7 Plus is $0.40 in and $1.60 out in Singapore against $0.276 and $1.101 elsewhere, 1.45×.
Recorded price changes
Every change this site has recorded for Qwen3.7-Plus, with the date it took or takes effect and why. Corrections of our own reading are listed too, and say so.
-
September 11, 2026 · change of this site's criterion
Change of criterion on our side, not a price change by Alibaba: no Alibaba price rose. Alibaba Cloud's Model Studio pages list the same model at two price levels: the China (Beijing) region and the Global regions (Frankfurt, Virginia, Tokyo, Hong Kong) share one level, and the Singapore region, labelled by Alibaba "Scope: International", is higher. Until September 11, 2026 the first-party table quoted the Beijing/Global level while the provider comparison quoted Alibaba's OpenRouter endpoint at the Singapore level, so the site showed two different "Alibaba" prices for the same model. From September 11 the reference region is Singapore (International), the one Alibaba labels international, the one every reseller we track (OpenRouter, Together, Fireworks) mirrors, and the higher of the two; the Beijing/Global figures stay published next to it. Published figures move from 1.65 / 4.951 to 2.00 / 6.00 (Qwen3.8 Max), 0.113 / 0.382 to 0.15 / 0.47 (Qwen3.8 Flash), 0.276 / 1.101 to 0.40 / 1.60 (Qwen3.7 Plus, prompts up to 256K; the tier above 256K, 1.20 / 4.80, is now recorded too) and 1.65 / 4.951 to 2.50 / 7.50 (Qwen3.7 Max). Alibaba's pages state that they show the original pricing excluding any limited-time promotions.
Every price change recorded on this site →
Specifications
| API model id | qwen3.7-plus | |
|---|---|---|
| Provider | Alibaba (Qwen) | |
| Context window | 1M tokens | |
| Max output | 131K tokens | |
| Released | May 26, 2026 | |
| Status | Current | |
Qwen3.7-Plus was released on May 26, 2026. It returns up to 131K tokens per response, against a context window of 1M tokens.
Questions about Qwen3.7-Plus pricing
- How much does Qwen3.7-Plus cost per million tokens?
- $0.40 per 1M input tokens and $1.60 per 1M output tokens in the Singapore (International) region for prompts up to 256,000 tokens, read on Alibaba (Qwen)'s documentation page on September 11, 2026.
- What does a cached input token cost on Qwen3.7-Plus?
- $0.08 per 1M cached input tokens, 20% of the $0.40 input price.
- Does Qwen3.7-Plus cost more for long prompts?
- Yes. For prompts over 256,000 tokens it costs $1.20 per 1M input tokens and $4.80 per 1M output tokens, $0.24 for cached input, against $0.40 and $1.60 below that.
Verified September 11, 2026 against Alibaba (Qwen)'s documentation page. Confidence: confirmed. How the data is checked.
Other Alibaba (Qwen) models with a verified price
- Qwen3.7-Flash $0.03 in · $0.13 out
- Qwen3.8-Flash $0.15 in · $0.47 out
- Qwen3.8-Omni-Flash $0.15 in · $0.47 out
- Qwen3.8-Max $2.00 in · $6.00 out
- Qwen3.7-Max $2.50 in · $7.50 out
Cheaper or equally priced alternatives
Current models with an input price at or below $0.40 per 1M tokens.
- Gemini 3.5 Flash-Lite Google $0.30 in · $2.50 out
- Gemini 2.5 Flash Google $0.30 in · $2.50 out
- Codestral Mistral $0.30 in · $0.90 out
- DeepSeek V4.1 Flash DeepSeek $0.30 in · $1.20 out
Compare every model in one table →
Work out what this costs at your volume →