Inference providers · open weights · Moonshot AI
Kimi K3
Open-weights model published by Moonshot AI, which also serves it through its own endpoint on OpenRouter: that rate is the maker's list price, and the other 3 hosts below are compared with it. Each figure was read on that host's own pricing page and links to it, with the date it was checked.
What each host charges
4 hosts, from $2.85 (DeepInfra) to $3.00 (OpenRouter) per 1M input tokens: a 1.1× spread, within the 1.15× that this site treats as the same price. Pick a host on latency, region or tooling; the bill will be the same.
| Host | Input | Output | Cached input | Context served | Checked | Source |
|---|---|---|---|---|---|---|
| DeepInfra | $2.85 | $14.25 | — | 1.0M tokens | September 30, 2026 | deepinfra.com |
| Together AI | $3.00 | $15.00 | $0.30 | 1.0M tokens | October 1, 2026 | together.ai |
| Fireworks AI | $3.00 | $15.00 | $0.30 | 1.0M tokens | October 1, 2026 | docs.fireworks.ai |
| OpenRouter (via Moonshot AI, the maker) | $3.00 | $15.00 | $0.30 | 1.0M tokens | October 1, 2026 | openrouter.ai |
DeepInfra charges $2.85 per 1M input tokens and OpenRouter $3.00 for the same weights; on output, $14.25 against $15.00.
Recorded price changes
Every change this site has recorded for a host's price of Kimi K3, with the date it took or takes effect and why. Changes of our own rule are listed too, and say so.
-
September 1, 2026 · Fireworks AI (all models on this host) · price up 1.5×
Since September 1, 2026 Fireworks adds a 50% surcharge to the base price of the Serverless models it marks as US-only. The Fireworks prices on this site are the base prices, without the surcharge.
On OpenRouter
On September 15, 2026 OpenRouter listed 20 endpoints for Kimi K3, 18 of them healthy, from $2.10 (InferenceNet) to $6.00 per 1M input tokens. The figure OpenRouter's catalog shows for the model, $2.6481 / $13.2827 (Sail Research), is the 6th cheapest of those 20. Of these, only Moonshot AI's own endpoint is the maker's list price; the rest are resellers. DeepInfra, Together AI and Fireworks AI, priced above from their own pages, are among them.
Why the catalog figure is not what a request pays →
Specifications
| Weights published by | Moonshot AI |
|---|---|
| Context length (maker) | 1.0M tokens |
| Context served by hosts | 1.0M tokens |
| API model ids | DeepInfra: moonshotai/Kimi-K3; Together AI: moonshotai/Kimi-K3; Fireworks AI: kimi-k3; OpenRouter: moonshotai/kimi-k3 |
| Hosts with a verified price | 4 |
Host prices verified between September 30, 2026 and October 1, 2026, each on the host's own pricing page linked in the table above, never from a reseller catalog or a third-party table. Maker and context length from Moonshot AI's model card, checked September 15, 2026 (max_position_embeddings 1048576 in the model's config.json (text_config)). How the data is checked.
Every open model, every host, one table →
When comparing hosts is worth it, and when it is not →
What the OpenRouter price means, and what a request actually pays →