Gemini 3.6 Flash
Google's current workhorse model, released July 2026 as the practical successor to 3.5 Flash rather than the long-delayed 3.5 Pro. It improves on coding, knowledge work and multimodal handling while reducing token usage by up to 17%, which makes it cheaper to run than its predecessor on the same task even before the headline rate is compared.
Gemini 3.6 Flash is a multimodal AI model from Google. It costs $1.500 per million input tokens and $7.500 per million output tokens (blended $3.300/M), with a 1,050,000-token context window.
- Up to 17% fewer tokens per task than 3.5 Flash
- Coding and knowledge work
- Multimodal input including audio and video
- 1M-token context
- High-volume production traffic
- Multimodal document and media analysis
- Coding assistants at scale
- Cost-sensitive agent steps
Benchmarks
More from Google
See all 24 →Frequently asked questions
How much does Gemini 3.6 Flash cost?
Gemini 3.6 Flash costs $1.500 per million input tokens and $7.500 per million output tokens, for a blended reference rate of $3.300 per million tokens.
What is Gemini 3.6 Flash's context window?
Gemini 3.6 Flash supports up to 1,050,000 tokens of context in a single request.
What is Gemini 3.6 Flash best for?
Gemini 3.6 Flash is well suited to Up to 17% fewer tokens per task than 3.5 Flash, Coding and knowledge work and Multimodal input including audio and video.
Who makes Gemini 3.6 Flash?
Gemini 3.6 Flash is developed and served by Google. It was released in Jul 2026.