Claude Fable 5$22.000/MClaude Opus 5$11.000/MClaude Opus 4.8$11.000/MClaude Opus 4.7$11.000/MClaude Opus 4.6$11.000/MClaude Opus 4.5$33.000/MClaude Sonnet 3.7$6.600/MClaude Opus 3$33.000/MClaude 2.1$12.800/MClaude 2$12.800/MGPT-5.6 Sol$12.500/MGPT-5.6 Terra$5.000/MGPT-5.5$12.500/MGPT-5.2$5.425/MGPT-5.2-Codex$5.425/MGPT-5$3.875/MGPT-4.5$97.500/MGPT-4 Turbo Preview$16.000/MGPT-4$39.000/MGPT-4-32k$78.000/Mo3$19.000/Mo3-mini$2.090/Mo4-mini$2.090/Mo1$28.500/Mo1-mini$5.700/Mo1-preview$28.500/MGemini 3.5 Pro$5.000/MGemini 3.1 Pro$5.000/MGemini 3 Pro$5.000/MGemini 2.5 Pro$3.875/MClaude Fable 5$22.000/MClaude Opus 5$11.000/MClaude Opus 4.8$11.000/MClaude Opus 4.7$11.000/MClaude Opus 4.6$11.000/MClaude Opus 4.5$33.000/MClaude Sonnet 3.7$6.600/MClaude Opus 3$33.000/MClaude 2.1$12.800/MClaude 2$12.800/MGPT-5.6 Sol$12.500/MGPT-5.6 Terra$5.000/MGPT-5.5$12.500/MGPT-5.2$5.425/MGPT-5.2-Codex$5.425/MGPT-5$3.875/MGPT-4.5$97.500/MGPT-4 Turbo Preview$16.000/MGPT-4$39.000/MGPT-4-32k$78.000/Mo3$19.000/Mo3-mini$2.090/Mo4-mini$2.090/Mo1$28.500/Mo1-mini$5.700/Mo1-preview$28.500/MGemini 3.5 Pro$5.000/MGemini 3.1 Pro$5.000/MGemini 3 Pro$5.000/MGemini 2.5 Pro$3.875/M
Google
Google
Multimodal

Gemini 3.6 Flash

Balanced

Google's current workhorse model, released July 2026 as the practical successor to 3.5 Flash rather than the long-delayed 3.5 Pro. It improves on coding, knowledge work and multimodal handling while reducing token usage by up to 17%, which makes it cheaper to run than its predecessor on the same task even before the headline rate is compared.

Gemini 3.6 Flash is a multimodal AI model from Google. It costs $1.500 per million input tokens and $7.500 per million output tokens (blended $3.300/M), with a 1,050,000-token context window.

INPUT
$1.500/M
per million input tokens
OUTPUT
$7.500/M
per million output tokens
CONTEXT
1,050,000
tokens
What it is good at
  • Up to 17% fewer tokens per task than 3.5 Flash
  • Coding and knowledge work
  • Multimodal input including audio and video
  • 1M-token context
Typical use cases
  • High-volume production traffic
  • Multimodal document and media analysis
  • Coding assistants at scale
  • Cost-sensitive agent steps

Benchmarks

No published benchmark scores tracked for this model yet. Frontier and reasoning models from major providers have scores; smaller models and inference-host variants typically inherit the underlying open-weights score.

More from Google

See all 24

Frequently asked questions

How much does Gemini 3.6 Flash cost?

Gemini 3.6 Flash costs $1.500 per million input tokens and $7.500 per million output tokens, for a blended reference rate of $3.300 per million tokens.

What is Gemini 3.6 Flash's context window?

Gemini 3.6 Flash supports up to 1,050,000 tokens of context in a single request.

What is Gemini 3.6 Flash best for?

Gemini 3.6 Flash is well suited to Up to 17% fewer tokens per task than 3.5 Flash, Coding and knowledge work and Multimodal input including audio and video.

Who makes Gemini 3.6 Flash?

Gemini 3.6 Flash is developed and served by Google. It was released in Jul 2026.

Compared with

Terms used on this page