Back to Blog

Gemini 3.8 Flash Listed: $0.75 Input, 1049k Context, Multi-Modal

EzraSeptember 3, 20261 min read
Gemini 3.8 Flash Listed: $0.75 Input, 1049k Context, Multi-Modal

Google Adds Gemini 3.8 Flash to the Roster

Google listed Gemini 3.8 Flash on 2026-09-02, slotting it into the mid-tier Flash family with a notably wide context window and multi-modal input support.

Pricing and Context at a Glance

  • Input: $0.75 per 1M tokens
  • Output: $3.75 per 1M tokens
  • Context window: 1049k tokens

The 5:1 output-to-input price ratio is fairly standard for Flash-class models, but the 1049k context window is worth flagging — it's large enough to handle lengthy codebases, legal documents, or extended conversation histories in a single pass without chunking workarounds.

Modalities

The model accepts text, image, video, file, and audio inputs and returns text. That breadth makes it relevant for transcription pipelines, document Q&A, and mixed-media workflows where swapping in a specialist model per modality would add latency and cost.

Who Should Care

Developers running long-context tasks get a competitive window without moving to a full "Pro" tier price. At $0.75 input, processing a 1M-token document costs $0.75 — manageable for batch jobs, tighter for real-time applications where output volume drives the bill.

Teams already on Gemini 2.x Flash should compare throughput and quality on their actual workloads before migrating; pricing alone doesn't tell the whole story.

Users of competing mid-range models (Claude Haiku, GPT-4o Mini, etc.) now have another data point — check our comparison tables to see where 3.8 Flash lands on price-per-output-token across the board.

Full pricing details, availability status, and historical price changes are tracked on the Gemini 3.8 Flash — live specs & price history page.

Ezra, Scout AI Team

E

Ezra

Ezra tracks the AI model market for the Scout AI Team — token prices, benchmarks and usage data from our live six-hour sync pipeline.