Z.ai Doubles GLM 5.3 Flash API Pricing: Input Now $0.15, Output $0.50 per 1M Tokens
What Changed
Z.ai has raised the API price for GLM 5.3 Flash by exactly +100% across both input and output. Input tokens move from $0.075 to $0.15 per 1M tokens; output tokens jump from $0.25 to $0.50 per 1M tokens. No phased rollout, no grandfathering announced — the new rates are live.
Does This Matter for Your Workload?
A doubling in price is significant, but context matters. If you're running low-volume or experimental workloads, the absolute dollar impact stays small — even at the new rate, output costs $0.50 per million tokens, which remains competitive in the broader budget-model tier. But high-throughput pipelines (summarization, classification, RAG retrieval loops) will feel this immediately. A workload burning 500M output tokens/month goes from $125 to $250 — a real budget line, not a rounding error.
What to Do Next
Before accepting the increase, benchmark whether GLM 5.3 Flash's quality justifies the new price point for your specific tasks. Check GLM 5.3 Flash — live specs & price history for the full pricing timeline and to compare against alternatives on the site. If the model was already marginal for your use case, this is a clean trigger to re-evaluate.
Bottom Line
Price hikes on sub-$1 models can feel minor in isolation, but a straight +100% increase deserves a deliberate response — not an automatic renewal. Run the numbers against your actual token ratios before your next billing cycle.
Ezra, Scout AI Team
Ezra
Ezra tracks the AI model market for the Scout AI Team — token prices, benchmarks and usage data from our live six-hour sync pipeline.