GLM (Z.ai) Pricing: Which Model Is Worth Paying For?
OpenRouter lists 16 GLM (Z.ai) models. For 1M input plus 1M output tokens they range from $0.46 (GLM 4.7 Flash) to $11.60 (GLM 5.3 Prime), a 25x spread. Most teams overpay by defaulting to the flagship for work a cheaper model passes.
Plans and pricing
GLM 4.7 Flash
$0.06 input / $0.40 output per 1M tokens200K context on OpenRouter (z-ai/glm-4.7-flash). 1M input plus 1M output tokens: $0.46.
Best for
High-volume, low-stakes steps: routing, tagging, extraction and first drafts.
Not for
Tasks where a wrong answer costs more than the tokens saved.
GLM 4.5
$0.60 input / $2.20 output per 1M tokens131K context on OpenRouter (z-ai/glm-4.5). 1M input plus 1M output tokens: $2.80.
Best for
A middle tier to test before paying flagship prices.
Not for
Tasks where a wrong answer costs more than the tokens saved.
GLM 5.3 Flash
$0.15 input / $0.50 output per 1M tokens1M context on OpenRouter (z-ai/glm-5.3-flash). 1M input plus 1M output tokens: $0.65.
Best for
A middle tier to test before paying flagship prices.
Not for
Tasks where a wrong answer costs more than the tokens saved.
GLM 5.3 Prime
$2.80 input / $8.80 output per 1M tokens1M context on OpenRouter (z-ai/glm-5.3-prime). 1M input plus 1M output tokens: $11.60.
Best for
The hardest GLM (Z.ai) tasks, once cheaper GLM (Z.ai) models have failed your acceptance tests.
Not for
Bulk work a smaller model already passes; the price gap compounds at volume.
Prices are OpenRouter's listed per-token prices for Z.ai (Zhipu) models, read from the OpenRouter models API and checked 2026-10-04. Buying direct from Z.ai (Zhipu) or another host can be priced differently; prices exclude tax and change often.
Who should pay for GLM (Z.ai)?
GLM (Z.ai) suits coding and agent workloads that need long context at a fraction of frontier prices. Start with GLM 4.7 Flash on a sample of real tasks, measure the acceptance rate, and step up a tier only where it fails. Price the whole job, not the token: retries and human review often cost more than the model.
Calculate your exact cost →Who is overpaying for GLM (Z.ai)?
- Teams sending every request to GLM 5.3 Prime when GLM 4.7 Flash passes most of them
- Anyone paying for output tokens they never read: cap max tokens and ask for concise answers
- Buyers comparing token prices without counting retries, failed responses and review time
Official pricing sources
Frequently asked questions
What is the cheapest GLM (Z.ai) model?
On OpenRouter, GLM 4.7 Flash at $0.06 input and $0.40 output per 1M tokens (checked 2026-10-04).
How much does the top GLM (Z.ai) model cost?
GLM 5.3 Prime is listed at $2.80 input and $8.80 output per 1M tokens, so 1M of each costs $11.60.
Is GLM (Z.ai) cheaper than GPT or Claude?
Usually per token, yes. Compare on your own prompts with the AI Cost Calculator: the cheaper model only saves money if it passes the same acceptance tests.