Pricing · DeepSeek
DeepSeek V4.1 Flash
DeepSeek's deepseek-flash, 1M tokens of context. Every number on this page was read off DeepSeek's own pricing
page — never estimated, never recalled from memory.
Pricing
| Input | $0.30 | per 1M tokens |
|---|---|---|
| Output | $1.20 | per 1M tokens |
| Cached input | $0.006 | per 1M tokens · 0.02× the input price |
| Off-peak input | $0.15 | per 1M tokens · outside Monday to Friday 01:00-04:00 and 06:00-10:00 UTC |
| Off-peak output | $0.60 | per 1M tokens · cached input $0.003 |
The headline prices are DeepSeek's peak rate, which applies only Monday to Friday 01:00-04:00 and 06:00-10:00 UTC: 35 of the 168 hours in a week (21%). The rest of the time every rate is the off-peak one above. Source: the same pricing page, checked September 14, 2026.
At the calculator's starting volume, 1M input tokens and 200K output tokens with no cache hits, DeepSeek V4.1 Flash costs $0.54: $0.30 for input and $0.24 for output, at the peak prices above. Same formula as the calculator; change the volume there.
Where this price sits
$0.30 per 1M input tokens is the 13th cheapest input price of the 67 current models with a verified price on this site, and the 1st of DeepSeek's 2. Output costs 4.0× the input price.
Mentioned in
2 posts on this site cite DeepSeek V4.1 Flash. The sentence is the post's own, taken from where it links here.
-
Prompt Caching Is Not 10% Everywhere: What a Cache Hit Costs Across Eight Providers · September 16, 2026
DeepSeek V4 Pro charges $0.044 per million cached tokens on $1.32 of input, and DeepSeek V4.1 Flash $0.006 on $0.30.
-
Five Things Decide What You Pay for a Model, and Only One of Them Is the Model · September 15, 2026
The same clock applies to DeepSeek V4.1 Flash: $0.30 and $1.20 at peak, $0.15 and $0.60 off-peak.
Succession
DeepSeek names DeepSeek V4.1 Flash as the replacement for:
- DeepSeek V4 Flash Vision (experimental) retired shut down September 10, 2026
- DeepSeek V4 Flash retired shut down September 10, 2026
Specifications
| API model id | deepseek-flash | |
|---|---|---|
| Provider | DeepSeek | |
| Context window | 1M tokens | |
| Max output | 384K tokens | |
| Released | September 10, 2026 | |
| Status | Current | |
DeepSeek V4.1 Flash was released on September 10, 2026. It returns up to 384K tokens per response, against a context window of 1M tokens.
Questions about DeepSeek V4.1 Flash pricing
- How much does DeepSeek V4.1 Flash cost per million tokens?
- $0.30 per 1M input tokens and $1.20 per 1M output tokens, read on DeepSeek's official pricing page on October 1, 2026.
- What does a cached input token cost on DeepSeek V4.1 Flash?
- $0.006 per 1M cached input tokens, 2% of the $0.30 input price.
Released September 10, 2026 with native vision; it replaced DeepSeek V4 Flash and V4 Flash Vision, whose old model names are still accepted and routed to this model at this price.
Verified October 1, 2026 against DeepSeek's official pricing page. Confidence: confirmed. No price change recorded for this model. How the data is checked.
Other DeepSeek models with a verified price
- DeepSeek V4 Pro $1.32 in · $3.96 out
Cheaper or equally priced alternatives
Current models with an input price at or below $0.30 per 1M tokens.
- Gemini 3.5 Flash-Lite Google $0.30 in · $2.50 out
- Gemini 2.5 Flash Google $0.30 in · $2.50 out
- Codestral Mistral $0.30 in · $0.90 out
- GPT-5.6 Luna OpenAI $0.20 in · $1.20 out
Compare every model in one table →
Work out what this costs at your volume →