Z.ai GLM-5.2 API Pricing
Last updated:
GLM-5.2 costs $1.40 input / $4.40 output per 1M tokens. Z.ai also lists cached input at $0.26 per 1M tokens and a 1M-token context window. For coding-tool users, Z.ai sells a separate GLM Coding Plan subscription that includes GLM-5.2 for supported tools.
Z.ai API token prices
GLM-5.2 | Z.ai | Mid | 1M | $1.40 | $0.26 | $4.40 |
- GLM-5.2Z.aiMid
- Input
- $1.40
- Cached
- $0.26
- Output
- $4.40
Sponsored links may earn us a commission at no extra cost to you. Affiliate status never changes model ordering.
All product names, logos, and brands are property of their respective owners and are used for identification purposes only.
If you are comparing GLM against other open models, Novita is a managed multi-model API to quote alongside first-party and hosted options. For very steady volume, check the GPU break-even guide.
Affiliate disclosure: this sponsored link may earn us a commission. It does not affect Z.ai table order or pricing data.
API vs GLM Coding Plan
Use the API when you are building software, agents, automations, customer-facing features, or anything outside Z.ai's supported coding-tool flow. Use the GLM Coding Plan when one developer wants high-volume GLM access inside Claude Code, OpenClaw, Cline, Kilo Code, OpenCode, Crush, Goose, or another supported coding tool.
The important catch: Z.ai says the GLM Coding Plan is strictly limited to supported tools and products. Calls outside the plan use normal API billing, and GLM-5.2 consumes more quota than GLM-4.7 during peak windows. Z.ai's DevPack docs list GLM-5.2 and GLM-5-Turbo at 3x quota during 14:00-18:00 UTC+8, 2x off-peak, and a limited-time 1x off-peak benefit through the end of September.
Frequently asked questions
How much does GLM-5.2 cost on the Z.ai API?
Z.ai lists GLM-5.2 at $1.40 per 1M input tokens, $0.26 per 1M cached input tokens, and $4.40 per 1M output tokens. Cached input storage is listed as limited-time free.
Does the GLM Coding Plan include GLM-5.2?
Yes. Z.ai says all GLM Coding Plan tiers support GLM-5.2, GLM-5-Turbo, GLM-4.7, and GLM-4.5-Air. The plan is limited to officially supported coding tools and is not a general API subscription.
Is GLM-5.2 cheaper through the subscription or API?
For supported coding tools, the subscription can be much cheaper because Z.ai describes the monthly quota as equivalent to roughly 15-30x the monthly subscription fee at API prices. For custom software, agents, SDK usage, or production apps, use the API pricing instead.
How does GLM Coding Plan quota affect GLM-5.2?
Z.ai says GLM-5.2 and GLM-5-Turbo consume 3x quota during the 14:00-18:00 UTC+8 peak window and 2x quota off-peak, with a limited-time 1x off-peak benefit through the end of September. Each prompt may invoke the model 15-20 times, so heavy repositories can burn quota faster than the headline prompt count suggests.