Back to Blog

Google Lists Gemini 3.6 Flash: $1.50 Input, 1049k Context Window

EzraJuly 22, 20261 min read
Google Lists Gemini 3.6 Flash: $1.50 Input, 1049k Context Window

What's New

Google added Gemini 3.6 Flash to its model lineup on 2026-07-21. It sits in the Flash tier — Google's cost-efficiency line — and supports a broad input mix: text, image, video, file, and audio, all resolving to text output.

Pricing & Context

At $1.50 per 1M input tokens and $7.50 per 1M output tokens, Gemini 3.6 Flash is priced like a workhorse model rather than a premium one. The 5:1 output-to-input price ratio is steep compared to some rivals, so workloads that generate long responses will feel that gap. Keep generation lengths in check if cost is the primary concern.

The 1049k-token context window is the headline spec here. That's enough to stuff in large codebases, lengthy document collections, or extended conversation histories without chunking. For RAG-light use cases — where you'd rather just dump the full document in — this is a practical advantage.

Who Should Care

  • Developers running long-document pipelines: The 1049k context reduces preprocessing overhead significantly.
  • Multimodal apps: Native support for image, video, audio, and file inputs in a single model call keeps architecture simple.
  • Cost-sensitive teams: $1.50 input is accessible, but watch output volume — at $7.50 per 1M tokens, a chatty pipeline adds up fast.

Compare it against alternatives on our site before committing: Gemini 3.6 Flash — live specs & price history.

Bottom Line

Gemini 3.6 Flash looks like a solid mid-tier option for multimodal, long-context tasks. The pricing is reasonable on the input side; output costs are the lever worth watching.

Ezra, Scout AI Team

E

Ezra

Ezra tracks the AI model market for the Scout AI Team — token prices, benchmarks and usage data from our live six-hour sync pipeline.