Gemini 3.5 Flash
The Flash tier of the Gemini 3.5 generation, a fast, multimodal workhorse with a 1M-token context window. It has since been succeeded by Gemini 3.6 Flash, which does the same job using fewer tokens, so 3.5 Flash is mainly of interest to workloads already tuned against it.
Gemini 3.5 Flash is a multimodal AI model from Google. It costs $1.500 per million input tokens and $9.000 per million output tokens (blended $3.750/M), with a 1,000,000-token context window.
- Fast multimodal inference
- 1M-token context
- Audio and video input
- Established behaviour for existing pipelines
- Pipelines already tuned to 3.5 Flash
- Multimodal extraction
- High-throughput classification
Benchmarks
More from Google
See all 24 →Frequently asked questions
How much does Gemini 3.5 Flash cost?
Gemini 3.5 Flash costs $1.500 per million input tokens and $9.000 per million output tokens, for a blended reference rate of $3.750 per million tokens.
What is Gemini 3.5 Flash's context window?
Gemini 3.5 Flash supports up to 1,000,000 tokens of context in a single request.
What is Gemini 3.5 Flash best for?
Gemini 3.5 Flash is well suited to Fast multimodal inference, 1M-token context and Audio and video input.
Who makes Gemini 3.5 Flash?
Gemini 3.5 Flash is developed and served by Google.