DeepSeek V4 Flash 0731 API Prices Double: Input $0.065, Output $0.18 per 1M Tokens
What Changed
DeepSeek has quietly doubled the API pricing on V4 Flash 0731. Input tokens moved from $0.045 to $0.065 per 1M tokens, while output tokens jumped from $0.09 to $0.18 per 1M tokens — a clean +100% increase on the output side.
Check the full pricing timeline on the DeepSeek V4 Flash 0731 — live specs & price history page.
Does This Matter for Your Workload?
It depends heavily on your input/output ratio. If your application is output-heavy — think long-form generation, code completion, or summarization with verbose responses — the doubling of output costs hits hardest. A workload burning 10M output tokens monthly just went from $0.90 to $1.80, which compounds fast at scale.
For read-heavy or classification tasks where output is minimal, the input bump from $0.045 to $0.065 is more modest in real terms.
Practical Takeaways
- Audit your token split. If you're outputting significantly more than you're inputting, recalculate your monthly bill immediately.
- Compare alternatives. This increase makes it worth benchmarking V4 Flash 0731 against other budget-tier models on the site — a 100% output price hike closes the gap with mid-tier options considerably.
- No announced capability change. This appears to be a pricing adjustment only; there's no indication of a model update or performance improvement accompanying the increase.
Pricing changes like this are a reminder to avoid locking architecture tightly to a single model endpoint without monitoring cost-per-token trends.
Ezra, Scout AI Team
Ezra
Ezra tracks the AI model market for the Scout AI Team — token prices, benchmarks and usage data from our live six-hour sync pipeline.