Z.ai GLM API pricing
Z.ai GLM API pricing: current per-token rates
Z.ai GLM's current API rates per 1M tokens, input / output / cached-read, verified 2026-07-23. Every rate links to its official source.
Current Z.ai GLM model prices
| Model | Input ($/1M) | Output ($/1M) | Cached in ($/1M) | Context |
|---|---|---|---|---|
| GLM 4.7 | $0.40 | $1.75 | $0.08 | 205K |
| GLM 5 | $0.95 | $3.15 | $0.19 | 205K |
| GLM 5.1 | $0.97 | $3.04 | $0.18 | 205K |
| GLM 5.2 | $0.80 | $2.52 | $0.15 | 1M |
Non-official data by OpenRouter top-provider pricing (open-weight model, hosted by multiple providers — the top provider’s rate, not necessarily the creator’s first-party price); verify against the provider's own page before relying on it.
Common questions
Can I use the Z.ai GLM API for free?
No — API access is metered per token at the rates above. Subscriptions (where offered) bundle usage into a flat monthly price; the API itself is pay-as-you-go.
Which model is cheapest?
The lowest input/output rate in the table above for light work; match the model to the task rather than defaulting to the cheapest — routing the hard steps up and the cheap steps down usually beats a single choice.
Prices change. Verified 2026-07-23 — re-check the source before relying on them. Corrections: hello@aiarch.dev.
Learn to cost-model agents like an architect.
What a Claude agent actually costs, or compare Professional Membership pricing.