Gemini 2.0 Flash
Efficient
Previous-generation Flash, still widely deployed. Strong default for cheap multimodal traffic.
Gemini 2.0 Flash is a efficient AI model from Google. It costs $0.100 per million input tokens and $0.400 per million output tokens (blended $0.190/M), with a 1,000,000-token context window.
INPUT
$0.100/M
per million input tokens
OUTPUT
$0.400/M
per million output tokens
CONTEXT
1,000,000
tokens
What it is good at
- Cheap multimodal
- 1M context
- Wide availability
Typical use cases
- Default cheap multimodal
- Workspace add-ons
Benchmarks
vs. best public score
Hand-curated from each provider's published reports and public leaderboards. Methodology varies across sources, treat as directional rather than authoritative.
More from Google
See all 24 →Gemini 3.5 Pro
Frontier · 1,000,000 ctx
in $2.000/Mout $12.000/M
Gemini 3.6 Flash
Balanced · 1,050,000 ctx
in $1.500/Mout $7.500/M
Gemini 3.5 Flash
Balanced · 1,000,000 ctx
in $1.500/Mout $9.000/M
Gemini 3.1 Pro
Frontier · 1,000,000 ctx
in $2.000/Mout $12.000/M
Gemini 3 Pro
Frontier · 1,000,000 ctx
in $2.000/Mout $12.000/M
Gemini 3 Flash
Efficient · 1,000,000 ctx
in $0.500/Mout $3.000/M
Frequently asked questions
How much does Gemini 2.0 Flash cost?
Gemini 2.0 Flash costs $0.100 per million input tokens and $0.400 per million output tokens, for a blended reference rate of $0.190 per million tokens.
What is Gemini 2.0 Flash's context window?
Gemini 2.0 Flash supports up to 1,000,000 tokens of context in a single request.
What is Gemini 2.0 Flash best for?
Gemini 2.0 Flash is well suited to Cheap multimodal, 1M context and Wide availability.
Who makes Gemini 2.0 Flash?
Gemini 2.0 Flash is developed and served by Google.