Inference providers · open weights · DeepSeek

DeepSeek V4 Flash 0731

Open-weights model published by DeepSeek; the API prices below are set by the 2 hosts that serve it, not by DeepSeek. There is no maker's list price to compare with: DeepSeek publishes the weights, not a per-token rate. Each figure was read on that host's own pricing page and links to it, with the date it was checked.

What each host charges

2 hosts, from $0.06 (DeepInfra) to $0.14 (Together AI) per 1M input tokens: a 2.3× spread. Above the 1.15× line this site draws, so the choice of host is a price decision.

Host Input Output Cached input Context served Checked Source
DeepInfra $0.06 $0.18 — 1.0M tokens September 30, 2026 deepinfra.com
Together AI $0.14 $0.28 $0.03 1.0M tokens October 1, 2026 together.ai
OpenRouter no list price: OpenRouter lists resellers only, none of them DeepSeek (27 endpoints on September 15, 2026) September 9, 2026 openrouter.ai

DeepInfra charges $0.06 per 1M input tokens and Together AI $0.14 for the same weights; on output, $0.18 against $0.28.

Hosts that no longer price it

Offerings this site tracked and stopped listing, with the date and the reason. They are kept because nobody else records when a host drops a model.

Recorded price changes

Every change this site has recorded for a host's price of DeepSeek V4 Flash 0731, with the date it took or takes effect and why. Changes of our own rule are listed too, and say so.

On OpenRouter

On September 15, 2026 OpenRouter listed 27 endpoints for DeepSeek V4 Flash 0731, 24 of them healthy, from $0.04 (OpenInference) to $0.44 per 1M input tokens. The figure OpenRouter's catalog shows for the model, $0.06 / $0.12 (Relace), is the 3rd cheapest of those 27. None of them is DeepSeek: all are resellers, which is why this site shows no OpenRouter price for the model. DeepInfra and Together AI, priced above from their own pages, are among them. Fireworks AI, which this site stopped listing on September 30, 2026 because its own pricing page carries no price for the model, serves it once through OpenRouter; that is a reseller listing, not a published price, so the row stays out.

Specifications

Weights published by DeepSeek
Context length (maker) 1.0M tokens
Context served by hosts 1.0M tokens
API model ids DeepInfra: deepseek-ai/DeepSeek-V4-Flash-0731; Together AI: deepseek-ai/DeepSeek-V4-Flash-0731
Hosts with a verified price 2

Host prices verified between September 30, 2026 and October 1, 2026, each on the host's own pricing page linked in the table above, never from a reseller catalog or a third-party table. Maker and context length from DeepSeek's model card, checked September 15, 2026 (max_position_embeddings 1048576 in the model's config.json). How the data is checked.