Back to Blog

Z.ai Doubles GLM 5.3 Flash API Pricing: Input Now $0.15, Output $0.50 per 1M Tokens

EzraSeptember 11, 20261 min read
Z.ai Doubles GLM 5.3 Flash API Pricing: Input Now $0.15, Output $0.50 per 1M Tokens

What Changed

Z.ai has raised the API price for GLM 5.3 Flash by exactly +100% across both input and output. Input tokens move from $0.075 to $0.15 per 1M tokens; output tokens jump from $0.25 to $0.50 per 1M tokens. No phased rollout, no grandfathering announced — the new rates are live.

Does This Matter for Your Workload?

A doubling in price is significant, but context matters. If you're running low-volume or experimental workloads, the absolute dollar impact stays small — even at the new rate, output costs $0.50 per million tokens, which remains competitive in the broader budget-model tier. But high-throughput pipelines (summarization, classification, RAG retrieval loops) will feel this immediately. A workload burning 500M output tokens/month goes from $125 to $250 — a real budget line, not a rounding error.

What to Do Next

Before accepting the increase, benchmark whether GLM 5.3 Flash's quality justifies the new price point for your specific tasks. Check GLM 5.3 Flash — live specs & price history for the full pricing timeline and to compare against alternatives on the site. If the model was already marginal for your use case, this is a clean trigger to re-evaluate.

Bottom Line

Price hikes on sub-$1 models can feel minor in isolation, but a straight +100% increase deserves a deliberate response — not an automatic renewal. Run the numbers against your actual token ratios before your next billing cycle.

Ezra, Scout AI Team

E

Ezra

Ezra tracks the AI model market for the Scout AI Team — token prices, benchmarks and usage data from our live six-hour sync pipeline.

Z.ai Doubles GLM 5.3 Flash API Pricing: Input Now $0.15, Output $0.50 per 1M Tokens | AIToolScout