Gemini 3.6 Flash makes the everyday model harder to dismiss.
Google's new Flash model targets coding, knowledge work and multimodal tasks with a one million token input window and a stronger focus on token efficiency.

Bit’s takeaway
What changed
Google released Gemini 3.6 Flash with text, image, video, audio and PDF input, a one million token input window, 64,000 output tokens, search and computer-use tools, and published API pricing of US$1.50 per million input tokens and US$7.50 per million output tokens.
Why it matters
The useful question is no longer whether a Flash model is the strongest model in general. It is whether this faster route now passes work that teams automatically send to a more expensive frontier model, especially repeated coding, document and multimodal tasks.
Who should care
- Teams routing a high volume of coding or knowledge work
- People working across long documents, images, audio or video
What to do
Take ten recurring tasks currently sent to a frontier model, run them through 3.6 Flash with the same inputs and score accuracy, correction effort, latency and cost per accepted result.
The human take
Tools change fast. Your judgment matters more.
Test it on one real, reviewable task before you trust it with anything that matters. A faster result is not automatically a better one.
Affected guidance
Put this to work
Sources and method