Z.ai Raises GLM 5.2 API Prices by 17% — What It Costs You Now
The Numbers
Z.ai has quietly pushed through a +17% price increase on its GLM 5.2 API. The new rates:
- Input: $0.554 → $0.65 per 1M tokens
- Output: $1.74 → $2.04 per 1M tokens
No new features or capability bump has been announced alongside the change — this is a straight cost increase.
Does It Actually Matter for Your Workload?
That depends heavily on your input/output ratio. Most real-world workloads are output-heavy, so the $0.30 per 1M output jump tends to sting more than the $0.096 input rise. Run a quick sanity check: take your monthly output token volume, divide by 1M, and multiply by $0.30 — that's your extra monthly spend from this change alone.
For low-volume or input-heavy use cases (batch classification, embeddings pipelines), the impact is modest. For high-volume chat or generation workloads, the compounding effect matters enough to revisit your model choice.
Worth Reconsidering Your Stack?
If GLM 5.2 wasn't a clear first choice before, a 17% hike is a reasonable prompt to re-evaluate. Several competing models in a similar capability tier are holding steadier pricing right now — use our comparison tools to benchmark cost per workload before committing.
Full rate details, historical pricing, and side-by-side comparisons are on the GLM 5.2 — live specs & price history page.
Bottom Line
No dramatic alarm bells, but no reason to ignore it either. If GLM 5.2 is a significant line item in your API budget, model the new output cost at $2.04 per 1M tokens and compare before your next billing cycle.
Ezra, Scout AI Team
Ezra
Ezra tracks the AI model market for the Scout AI Team — token prices, benchmarks and usage data from our live six-hour sync pipeline.