Gemini 2.5 Flash API Price Doubles: What Developers Need to Know
The Change
Google has doubled the API price of Gemini 2.5 Flash across the board. Input tokens move from $0.375 to $0.75 per 1M tokens, and output tokens jump from $1.88 to $3.75 per 1M tokens — a clean +100% increase on both lines.
No new capabilities, no new context window, no new benchmark scores announced alongside it. This is a straight price adjustment.
Does It Matter for Your Workload?
That depends heavily on your output-to-input ratio and volume.
- Low-volume or internal tooling: At these absolute numbers, even post-increase costs are modest. A million output tokens at $3.75 is still cheap for occasional use. This tier of developer probably won't notice.
- High-volume production apps: A 100% increase compounds fast. If you were running 100M output tokens a month at $1.88, that bill just became $375 instead of $188. Budget accordingly.
- Output-heavy workloads (summarisation, code generation, long-form drafts) feel the increase most — output pricing doubled just like input, but output volume typically drives the majority of cost.
Worth Comparing Alternatives
If Gemini 2.5 Flash was in your stack primarily because of price, now is a good time to re-benchmark against competing models. Speed, quality, and price have all shifted. Check the Gemini 3.7 Flash — live specs & price history page for a running timeline of this model's pricing and how it stacks up against alternatives tracked on this site.
Bottom Line
Double the price with no announced capability improvement is worth flagging. It doesn't make Flash unusable, but it does change the cost calculus — especially at scale. Re-run your numbers before assuming last month's estimate still holds.
Ezra, Scout AI Team
Ezra
Ezra tracks the AI model market for the Scout AI Team — token prices, benchmarks and usage data from our live six-hour sync pipeline.