Gemini 2.5 Flash
Balanced
Speed-optimised Gemini for high-throughput production. Same multimodal surface as Pro at a fraction of the price.
Gemini 2.5 Flash is a multimodal AI model from Google. It costs $0.150 per million input tokens and $0.600 per million output tokens (blended $0.285/M), with a 1,000,000-token context window.
INPUT
$0.150/M
per million input tokens
OUTPUT
$0.600/M
per million output tokens
CONTEXT
1,000,000
tokens
What it is good at
- Fast
- 1M context
- Native multimodal
- Very cheap
Typical use cases
- Production chat
- Bulk multimodal pipelines
- Cost-sensitive RAG
Benchmarks
vs. best public score
Hand-curated from each provider's published reports and public leaderboards. Methodology varies across sources, treat as directional rather than authoritative.
More from Google
See all 24 →Gemini 3.5 Pro
Frontier · 1,000,000 ctx
in $2.000/Mout $12.000/M
Gemini 3.6 Flash
Balanced · 1,050,000 ctx
in $1.500/Mout $7.500/M
Gemini 3.5 Flash
Balanced · 1,000,000 ctx
in $1.500/Mout $9.000/M
Gemini 3.1 Pro
Frontier · 1,000,000 ctx
in $2.000/Mout $12.000/M
Gemini 3 Pro
Frontier · 1,000,000 ctx
in $2.000/Mout $12.000/M
Gemini 3 Flash
Efficient · 1,000,000 ctx
in $0.500/Mout $3.000/M
Frequently asked questions
How much does Gemini 2.5 Flash cost?
Gemini 2.5 Flash costs $0.150 per million input tokens and $0.600 per million output tokens, for a blended reference rate of $0.285 per million tokens.
What is Gemini 2.5 Flash's context window?
Gemini 2.5 Flash supports up to 1,000,000 tokens of context in a single request.
What is Gemini 2.5 Flash best for?
Gemini 2.5 Flash is well suited to Fast, 1M context and Native multimodal.
Who makes Gemini 2.5 Flash?
Gemini 2.5 Flash is developed and served by Google.