Nemotron 3.5 Lightning API Prices Rise 25% — What It Costs You Now
What Changed
NVIDIA has raised API pricing on Nemotron 3.5 Lightning by +25% across the board. Input tokens move from $0.08 → $0.10 per 1M tokens, and output tokens from $0.20 → $0.25 per 1M tokens.
The ratio between input and output pricing stays the same (2.5×), so the cost structure hasn't fundamentally shifted — everything is just a quarter more expensive.
Does It Matter for Your Workload?
At these price points, Nemotron 3.5 Lightning remains in the budget tier, but the 25% jump is meaningful at scale. Some quick math:
- 10M input tokens/month: was $0.80, now $1.00
- 10M output tokens/month: was $2.00, now $2.50
- 100M mixed tokens/month: expect ~$25–35 in additional monthly spend depending on your input/output split
For low-volume hobby projects, this is noise. For production pipelines burning tens of millions of tokens monthly, it's worth re-evaluating your model choice — especially if you benchmarked cost-efficiency before the increase.
Worth Reconsidering Alternatives?
If you were on Nemotron 3.5 Lightning primarily for its price point, the +25% bump narrows its gap against competing models. Check the Nemotron 3.5 Lightning — live specs & price history page to compare it directly against alternatives ranked by cost-per-token and capability — filtering by your actual input/output ratio will give you the clearest picture.
Bottom Line
This isn't a dramatic repricing, but 25% is a real increase. If Nemotron 3.5 Lightning is a significant line item, now is a reasonable time to run a quick cost audit. If it's incidental spend, move on.
Ezra, Scout AI Team
Ezra
Ezra tracks the AI model market for the Scout AI Team — token prices, benchmarks and usage data from our live six-hour sync pipeline.