Pricing · Mistral
GLM 5.3
Mistral's zai-glm-5-3, 1M tokens of context. Every number on this page was read off Mistral's own pricing
page — never estimated, never recalled from memory.
How Mistral describes it
“Third-party model specializing in long-context agentic workflows and coding.”
Pricing
| Input | $1.40 | per 1M tokens |
|---|---|---|
| Output | $4.40 | per 1M tokens |
| Cached input | $0.14 | per 1M tokens · 0.1× the input price |
At the calculator's starting volume, 1M input tokens and 200K output tokens with no cache hits, GLM 5.3 costs $2.28: $1.40 for input and $0.88 for output, at the prices above. Same formula as the calculator; change the volume there.
Where this price sits
$1.40 per 1M input tokens is the 34th cheapest input price of the 67 current models with a verified price on this site, and the 7th of Mistral's 9. Output costs 3.1× the input price.
Mentioned in
One post on this site cites GLM 5.3. The sentence is the post's own, taken from where it links here.
-
Prompt Caching Is Not 10% Everywhere: What a Cache Hit Costs Across Eight Providers · September 16, 2026
Mistral's two GLM rows, GLM 5.2 and GLM 5.3, both $0.14 on $1.40, September 14 and 16.
Recorded price changes
Every change this site has recorded for GLM 5.3, with the date it took or takes effect and why. Corrections of our own reading are listed too, and say so.
-
September 16, 2026 · added to this dataset, our omission
Omission on our side, not a price change by any provider. On September 16, 2026, while reading OpenAI's pricing page by hand for a check on batch and flex prices, GPT-6 Astra turned up in the Standard table with no row in this dataset. A hand sweep of the first-party pricing pages of OpenAI, Anthropic, Google, xAI, Mistral, DeepSeek, Alibaba and Cohere against the dataset then found eight models charged per token with text output and a published price that had no row: GPT-6 Astra and GPT-Rosalind (OpenAI), Claude Haiku 3.5 (Anthropic, retired on the first-party API but still priced), Gemini 2.5 Computer Use Preview, Gemini Robotics ER 2 Preview and its Streaming variant, and Gemini Omni Flash Preview, a separate endpoint id from the generally available Omni Flash row at the same prices (Google), and GLM 5.3 (served by Mistral). Astra had been in OpenAI's Standard table at least since September 10, 2026, when the weekly detector first received that page's text, and this site's OpenRouter explainer already named it from the September 15 snapshot. The weekly detector could not have caught any of these: it compares the rows that exist against the page and has no step for models the page lists and the dataset does not. Each new row carries the page quote it was read from. Audio, image, video, embedding and per-minute models found in the same sweep were left out on purpose: the table covers models charged per token with text output.
Every price change recorded on this site →
As an open-weights model
GLM 5.3 is Z.ai's GLM-5.3, an open-weights model that Mistral serves through its own API at the price above. This site also tracks 3 other hosts that serve it, all at $1.40 per 1M input tokens (a spread of 1.0×, within the 1.15× this site treats as the same price), each with its date and source: GLM-5.3 across hosts →
Against GLM 5.2
GLM 5.2 is the earlier version of the same name in this dataset (still current). GLM 5.3 has the same prices as GLM 5.2: input price ($1.40), output price ($4.40), cached input price ($0.14) per 1M tokens. Its prices were verified September 28, 2026: GLM 5.2 pricing →
Specifications
| API model id | zai-glm-5-3 | |
|---|---|---|
| Provider | Mistral | |
| Family | GLM | |
| Context window | 1M tokens | |
| Max output | 128K tokens | |
| Status | Current | |
It returns up to 128K tokens per response, against a context window of 1M tokens.
Questions about GLM 5.3 pricing
- How much does GLM 5.3 cost per million tokens?
- $1.40 per 1M input tokens and $4.40 per 1M output tokens, read on Mistral's official pricing page on September 28, 2026.
- What does a cached input token cost on GLM 5.3?
- $0.14 per 1M cached input tokens, 10% of the $1.40 input price.
A Z.ai model served through Mistral's API. Added 2026-09-16 after a hand sweep of Mistral's pricing page found it missing from this dataset (our omission). Pricing page card: 'GLM 5.3 New Third-party model specializing in long-context agentic workflows and coding. Input (/M tokens) $1.4 Output (/M tokens) $4.4'. Cached input, context and max output from Mistral's model page docs.mistral.ai/models/zai-glm-5-3, read 2026-09-16: 'zai-glm-5-3 … Context 1M Max output 128k … $ 1.4 Input /M Tokens $ 0.14 Cached input /M Tokens $ 4.4 Output /M Tokens'.
Verified September 28, 2026 against Mistral's official pricing page. Confidence: confirmed. How the data is checked.
Other Mistral models with a verified price
- Ministral 3 3B $0.10 in · $0.10 out
- Mistral Small 4 $0.15 in · $0.60 out
- Ministral 3 8B $0.15 in · $0.15 out
- Ministral 3 14B $0.20 in · $0.20 out
- Codestral $0.30 in · $0.90 out
- Mistral Large 3 $0.50 in · $1.50 out
- GLM 5.2 $1.40 in · $4.40 out
- Mistral Medium 3.5 $1.50 in · $7.50 out
Cheaper or equally priced alternatives
Current models with an input price at or below $1.40 per 1M tokens.
- GLM 5.2 Mistral $1.40 in · $4.40 out
- DeepSeek V4 Pro DeepSeek $1.32 in · $3.96 out
- Gemini 2.5 Pro Google $1.25 in · $10.00 out
- Grok 4.3 xAI $1.25 in · $2.50 out
Compare every model in one table →
Work out what this costs at your volume →