Gemini 3 Flash
The efficient tier of the Gemini 3 generation: a 1M-token context window and multimodal input at a fraction of the Pro rate, aimed at high-volume work where latency and cost matter more than peak reasoning depth.
Gemini 3 Flash is a efficient AI model from Google. It costs $0.500 per million input tokens and $3.000 per million output tokens (blended $1.250/M), with a 1,000,000-token context window.
- Low cost per token
- 1M-token context
- Multimodal input
- Fast responses
- High-volume classification and extraction
- Routing and triage
- Cheap agent sub-steps
- Bulk summarisation
Benchmarks
More from Google
See all 24 →Frequently asked questions
How much does Gemini 3 Flash cost?
Gemini 3 Flash costs $0.500 per million input tokens and $3.000 per million output tokens, for a blended reference rate of $1.250 per million tokens.
What is Gemini 3 Flash's context window?
Gemini 3 Flash supports up to 1,000,000 tokens of context in a single request.
What is Gemini 3 Flash best for?
Gemini 3 Flash is well suited to Low cost per token, 1M-token context and Multimodal input.
Who makes Gemini 3 Flash?
Gemini 3 Flash is developed and served by Google.