Inference providers · open weights · Z.ai
GLM-5.3
Open-weights model published by Z.ai, which also serves it through its own endpoint on OpenRouter: that rate is the maker's list price, and the other 2 hosts below are compared with it. Each figure was read on that host's own pricing page and links to it, with the date it was checked.
What each host charges
3 hosts, every one of them at $1.40 per 1M input tokens and $4.40 per 1M output tokens: no spread at all, within the 1.15× that this site treats as the same price. Pick a host on latency, region or tooling; the bill will be the same.
| Host | Input | Output | Cached input | Context served | Checked | Source |
|---|---|---|---|---|---|---|
| Together AI | $1.40 | $4.40 | $0.26 | 1.0M tokens | October 1, 2026 | together.ai |
| Fireworks AI | $1.40 | $4.40 | $0.26 | — | October 1, 2026 | docs.fireworks.ai |
| OpenRouter (via Z.ai, the maker) | $1.40 | $4.40 | — | 1.3M tokens | October 1, 2026 | openrouter.ai |
Recorded price changes
Every change this site has recorded for a host's price of GLM-5.3, with the date it took or takes effect and why. Changes of our own rule are listed too, and say so.
-
September 1, 2026 · Fireworks AI (all models on this host) · price up 1.5×
Since September 1, 2026 Fireworks adds a 50% surcharge to the base price of the Serverless models it marks as US-only. The Fireworks prices on this site are the base prices, without the surcharge.
-
September 9, 2026 · OpenRouter · change of this site's criterion · figure shown here up 1.3×
The published OpenRouter price of GLM-5.3 rises 1.25x on both input ($1.12 to $1.40) and output ($3.52 to $4.40). No host raised its rate. The number changes because the rule changed: until September 9, 2026 this column showed the cheapest healthy endpoint at the moment of checking (GMICloud); from September 9 it shows the model maker's own price on OpenRouter's providers page, without promotional discounts (Z.ai, $1.40 / $4.40). Cheaper resellers still exist and are still reachable with :floor.
On OpenRouter
On September 15, 2026 OpenRouter listed 29 endpoints for GLM-5.3, 27 of them healthy, from $0.8775 (Reka) to $2.10 per 1M input tokens. The figure OpenRouter's catalog shows for the model, $1.40 / $4.40, is the 17th cheapest of those 29. Of these, only Z.ai's own endpoint is the maker's list price; the rest are resellers. Together AI and Fireworks AI, priced above from their own pages, are among them.
Why the catalog figure is not what a request pays →
Specifications
| Weights published by | Z.ai |
|---|---|
| Context length (maker) | 1.0M tokens |
| Context served by hosts | 1.0M tokens, 1.3M tokens |
| API model ids | Together AI: zai-org/GLM-5.3; Fireworks AI: glm-5p3; OpenRouter: z-ai/glm-5.3 |
| Hosts with a verified price | 3 |
Host prices verified on October 1, 2026, each on the host's own pricing page linked in the table above, never from a reseller catalog or a third-party table. Maker and context length from Z.ai's model card, checked September 15, 2026 (max_position_embeddings 1048576 in the model's config.json). How the data is checked.
Every open model, every host, one table →
When comparing hosts is worth it, and when it is not →
What the OpenRouter price means, and what a request actually pays →