Gemini 2.0 Flash Lite
Micro
Cheapest Gemini 2.0 tier. Targets very high-volume, latency-sensitive traffic where Flash is still too expensive.
Gemini 2.0 Flash Lite is a efficient AI model from Google. It costs $0.075 per million input tokens and $0.300 per million output tokens (blended $0.142/M), with a 1,000,000-token context window.
INPUT
$0.075/M
per million input tokens
OUTPUT
$0.300/M
per million output tokens
CONTEXT
1,000,000
tokens
What it is good at
- Cheapest Gemini
- Low latency
- 1M context
Typical use cases
- Bulk classification
- Realtime UX
- Cheap Workspace add-ons
Benchmarks
vs. best public score
Hand-curated from each provider's published reports and public leaderboards. Methodology varies across sources, treat as directional rather than authoritative.
More from Google
See all 24 →Gemini 3.5 Pro
Frontier · 1,000,000 ctx
in $2.000/Mout $12.000/M
Gemini 3.6 Flash
Balanced · 1,050,000 ctx
in $1.500/Mout $7.500/M
Gemini 3.5 Flash
Balanced · 1,000,000 ctx
in $1.500/Mout $9.000/M
Gemini 3.1 Pro
Frontier · 1,000,000 ctx
in $2.000/Mout $12.000/M
Gemini 3 Pro
Frontier · 1,000,000 ctx
in $2.000/Mout $12.000/M
Gemini 3 Flash
Efficient · 1,000,000 ctx
in $0.500/Mout $3.000/M
Frequently asked questions
How much does Gemini 2.0 Flash Lite cost?
Gemini 2.0 Flash Lite costs $0.075 per million input tokens and $0.300 per million output tokens, for a blended reference rate of $0.142 per million tokens.
What is Gemini 2.0 Flash Lite's context window?
Gemini 2.0 Flash Lite supports up to 1,000,000 tokens of context in a single request.
What is Gemini 2.0 Flash Lite best for?
Gemini 2.0 Flash Lite is well suited to Cheapest Gemini, Low latency and 1M context.
Who makes Gemini 2.0 Flash Lite?
Gemini 2.0 Flash Lite is developed and served by Google.