Back to Blog

Qwen3.5 397B A17B Drops Input Price 33% to $0.39 per 1M Tokens

EzraSeptember 7, 20261 min read
Qwen3.5 397B A17B Drops Input Price 33% to $0.39 per 1M Tokens

What Changed

Qwen has cut pricing on Qwen3.5 397B A17B across both token directions. Input drops from $0.55 to $0.39 per 1M tokens, and output falls from $3.50 to $2.34 per 1M tokens — a -33% reduction across the board.

Does It Move the Needle?

For read-heavy or retrieval-augmented workloads where input tokens dominate, the $0.39 input rate is genuinely competitive for a model at this parameter scale. The output cut matters more for generation-heavy tasks: at $2.34 per 1M output tokens, costs on a chatbot or long-form writing pipeline drop meaningfully at volume.

To illustrate: a workload burning 10M output tokens monthly just went from $35.00 to $23.40 — a $11.60 monthly saving before any input costs are counted.

Who Should Care

  • High-volume API users routing through Qwen3.5 397B A17B will see automatic savings with no integration changes required.
  • Teams evaluating MoE alternatives should reprice their benchmarks now — the gap between this model and some smaller, cheaper options has narrowed.
  • Cost-sensitive prototypers who previously ruled out a 397B-class model on price may want to revisit the numbers.

Check current rates, context limits, and how this model stacks up against alternatives on the Qwen3.5 397B A17B — live specs & price history.

Bottom Line

A -33% price cut on both input and output is a real reduction, not a rounding adjustment. If you're running Qwen3.5 397B A17B at any meaningful scale, your bill just got smaller automatically. If you weren't — it's worth a second look.

Ezra, Scout AI Team

E

Ezra

Ezra tracks the AI model market for the Scout AI Team — token prices, benchmarks and usage data from our live six-hour sync pipeline.