Z.ai GLM API pricing

Z.ai GLM API pricing: current per-token rates

Short answer

Z.ai GLM's current API rates per 1M tokens, input / output / cached-read, verified 2026-09-02. Every rate links to its official source.

Current Z.ai GLM model prices

ModelInput ($/1M)Output ($/1M)Cached in ($/1M)Context
GLM 4.7$0.40$1.75$0.08205K
GLM 5$0.60$1.92$0.12205K
GLM 5.1$0.97$3.04$0.18205K
GLM 5.2$0.97$3.04$0.191M

Non-official data by OpenRouter top-provider pricing (open-weight model, hosted by multiple providers — the top provider’s rate, not necessarily the creator’s first-party price); verify against the provider's own page before relying on it.

Recent price changes

  • GLM 5.2: $1.19/$3.74 → $0.97/$3.04 per 1M (effective 2026-09-02)
  • GLM 5.1: $1.26/$3.96 → $0.97/$3.04 per 1M (effective 2026-08-30)
  • GLM 5: $0.95/$2.55 → $0.60/$1.92 per 1M (effective 2026-08-15)

Common questions

Can I use the Z.ai GLM API for free?

No — API access is metered per token at the rates above. Subscriptions (where offered) bundle usage into a flat monthly price; the API itself is pay-as-you-go.

Which model is cheapest?

The lowest input/output rate in the table above for light work; match the model to the task rather than defaulting to the cheapest — routing the hard steps up and the cheap steps down usually beats a single choice.

Sources & provenance

Prices change. Verified 2026-09-02 — re-check the source before relying on them. Corrections: hello@aiarch.dev.