Back to Blog

Z.ai Raises GLM 5.2 API Prices by 17% — What It Costs You Now

EzraSeptember 20, 20261 min read
Z.ai Raises GLM 5.2 API Prices by 17% — What It Costs You Now

The Numbers

Z.ai has quietly pushed through a +17% price increase on its GLM 5.2 API. The new rates:

  • Input: $0.554 → $0.65 per 1M tokens
  • Output: $1.74 → $2.04 per 1M tokens

No new features or capability bump has been announced alongside the change — this is a straight cost increase.

Does It Actually Matter for Your Workload?

That depends heavily on your input/output ratio. Most real-world workloads are output-heavy, so the $0.30 per 1M output jump tends to sting more than the $0.096 input rise. Run a quick sanity check: take your monthly output token volume, divide by 1M, and multiply by $0.30 — that's your extra monthly spend from this change alone.

For low-volume or input-heavy use cases (batch classification, embeddings pipelines), the impact is modest. For high-volume chat or generation workloads, the compounding effect matters enough to revisit your model choice.

Worth Reconsidering Your Stack?

If GLM 5.2 wasn't a clear first choice before, a 17% hike is a reasonable prompt to re-evaluate. Several competing models in a similar capability tier are holding steadier pricing right now — use our comparison tools to benchmark cost per workload before committing.

Full rate details, historical pricing, and side-by-side comparisons are on the GLM 5.2 — live specs & price history page.

Bottom Line

No dramatic alarm bells, but no reason to ignore it either. If GLM 5.2 is a significant line item in your API budget, model the new output cost at $2.04 per 1M tokens and compare before your next billing cycle.

Ezra, Scout AI Team

E

Ezra

Ezra tracks the AI model market for the Scout AI Team — token prices, benchmarks and usage data from our live six-hour sync pipeline.