Inference providers · open weights · DeepSeek
DeepSeek V4 Flash 0731
Open-weights model published by DeepSeek; the API prices below are set by the 2 hosts that serve it, not by DeepSeek. There is no maker's list price to compare with: DeepSeek publishes the weights, not a per-token rate. Each figure was read on that host's own pricing page and links to it, with the date it was checked.
What each host charges
2 hosts, from $0.06 (DeepInfra) to $0.14 (Together AI) per 1M input tokens: a 2.3× spread. Above the 1.15× line this site draws, so the choice of host is a price decision.
| Host | Input | Output | Cached input | Context served | Checked | Source |
|---|---|---|---|---|---|---|
| DeepInfra | $0.06 | $0.18 | — | 1.0M tokens | September 30, 2026 | deepinfra.com |
| Together AI | $0.14 | $0.28 | $0.03 | 1.0M tokens | October 1, 2026 | together.ai |
| OpenRouter | no list price: OpenRouter lists resellers only, none of them DeepSeek (27 endpoints on September 15, 2026) | September 9, 2026 | openrouter.ai | |||
DeepInfra charges $0.06 per 1M input tokens and Together AI $0.14 for the same weights; on output, $0.18 against $0.28.
Hosts that no longer price it
Offerings this site tracked and stopped listing, with the date and the reason. They are kept because nobody else records when a host drops a model.
- Fireworks AI removed September 30, 2026 · last listed at $0.22 in · $0.66 out No longer listed on Fireworks' pricing page (docs.fireworks.ai/serverless/pricing), read on 2026-09-30. The page says models not listed individually are priced by size; this site does not know whether the model is still served.
Recorded price changes
Every change this site has recorded for a host's price of DeepSeek V4 Flash 0731, with the date it took or takes effect and why. Changes of our own rule are listed too, and say so.
-
September 30, 2026 · DeepInfra · price down 1.3×
DeepInfra lowered the input price of DeepSeek V4 Flash 0731 from $0.08 to $0.06 per million tokens; output stays at $0.18. It was $0.08 when this site read DeepInfra's pricing page on September 6, 2026 and $0.06 when it read it again on September 30, 2026; the exact day of the change is not known, so the date here is the day it was seen.
-
September 11, 2026 · OpenRouter · reference price withdrawn
DeepSeek's own endpoint for DeepSeek V4 Flash 0731 disappeared from OpenRouter between the afternoon of September 9, 2026, when it was still listed at $0.22 / $0.66 and this row was verified against it, and the 04:30 UTC check of September 10, which found no DeepSeek endpoint; it had not returned by September 11. DeepSeek still serves V4 Pro 0813 on OpenRouter, so this is specific to this model, not a general withdrawal. This site's OpenRouter column shows the model maker's list price, and for this model there is now none to show: the column moves from $0.22 / $0.66 per million tokens (last verified September 9) to "no list price". Nothing changed in our criterion, and no host changed a price; the provider withdrew its endpoint. The model is still served on OpenRouter by resellers (28 listed, 26 healthy on September 11), reachable with :floor; that day the catalog quoted $0.065 / $0.18 via Relace.
-
September 9, 2026 · OpenRouter · change of this site's criterion · figure shown here up 4.4×
The published OpenRouter price of DeepSeek V4 Flash 0731 rises 4.4x on input ($0.05 to $0.22) and 4.1x on output ($0.16 to $0.66). No host raised its rate. The number changes because the rule changed: until September 9, 2026 this column showed the cheapest healthy endpoint at the moment of checking (OpenInference); from September 9 it shows the model maker's own price on OpenRouter's providers page, without promotional discounts (DeepSeek, $0.22 / $0.66). Cheaper resellers still exist and are still reachable with :floor.
On OpenRouter
On September 15, 2026 OpenRouter listed 27 endpoints for DeepSeek V4 Flash 0731, 24 of them healthy, from $0.04 (OpenInference) to $0.44 per 1M input tokens. The figure OpenRouter's catalog shows for the model, $0.06 / $0.12 (Relace), is the 3rd cheapest of those 27. None of them is DeepSeek: all are resellers, which is why this site shows no OpenRouter price for the model. DeepInfra and Together AI, priced above from their own pages, are among them. Fireworks AI, which this site stopped listing on September 30, 2026 because its own pricing page carries no price for the model, serves it once through OpenRouter; that is a reseller listing, not a published price, so the row stays out.
Why the catalog figure is not what a request pays →
Specifications
| Weights published by | DeepSeek |
|---|---|
| Context length (maker) | 1.0M tokens |
| Context served by hosts | 1.0M tokens |
| API model ids | DeepInfra: deepseek-ai/DeepSeek-V4-Flash-0731; Together AI: deepseek-ai/DeepSeek-V4-Flash-0731 |
| Hosts with a verified price | 2 |
Host prices verified between September 30, 2026 and October 1, 2026, each on the host's own pricing page linked in the table above, never from a reseller catalog or a third-party table. Maker and context length from DeepSeek's model card, checked September 15, 2026 (max_position_embeddings 1048576 in the model's config.json). How the data is checked.
Every open model, every host, one table →
When comparing hosts is worth it, and when it is not →
What the OpenRouter price means, and what a request actually pays →