Z.ai GLM API pricing

Z.ai GLM API pricing: current per-token rates

Short answer

Z.ai GLM's current API rates per 1M tokens, input / output / cached-read, verified 2026-07-23. Every rate links to its official source.

Current Z.ai GLM model prices

ModelInput ($/1M)Output ($/1M)Cached in ($/1M)Context
GLM 4.7$0.40$1.75$0.08205K
GLM 5$0.95$3.15$0.19205K
GLM 5.1$0.97$3.04$0.18205K
GLM 5.2$0.80$2.52$0.151M

Non-official data by OpenRouter top-provider pricing (open-weight model, hosted by multiple providers — the top provider’s rate, not necessarily the creator’s first-party price); verify against the provider's own page before relying on it.

Common questions

Can I use the Z.ai GLM API for free?

No — API access is metered per token at the rates above. Subscriptions (where offered) bundle usage into a flat monthly price; the API itself is pay-as-you-go.

Which model is cheapest?

The lowest input/output rate in the table above for light work; match the model to the task rather than defaulting to the cheapest — routing the hard steps up and the cheap steps down usually beats a single choice.

Sources & provenance

Prices change. Verified 2026-07-23 — re-check the source before relying on them. Corrections: hello@aiarch.dev.