Gemini 2.5 Flash (batch)
Speed-optimised Gemini for high-throughput production. Same multimodal surface as Pro at a fraction of the price.
Gemini 2.5 Flash (batch) is a multimodal AI model from Google. It costs $0.150 per million input tokens and $1.250 per million output tokens (blended $0.480/M), with a 1,048,576-token context window.
Profile inherited from upstream Gemini 2.5 Flash ↗ — this is a hosted variant of the same open-weights model.
- Fast
- 1M context
- Native multimodal
- Very cheap
- Production chat
- Bulk multimodal pipelines
- Cost-sensitive RAG
Benchmarks
More from Google
See all 24 →Frequently asked questions
How much does Gemini 2.5 Flash (batch) cost?
Gemini 2.5 Flash (batch) costs $0.150 per million input tokens and $1.250 per million output tokens, for a blended reference rate of $0.480 per million tokens.
What is Gemini 2.5 Flash (batch)'s context window?
Gemini 2.5 Flash (batch) supports up to 1,048,576 tokens of context in a single request.
What is Gemini 2.5 Flash (batch) best for?
Gemini 2.5 Flash (batch) is well suited to Fast, 1M context and Native multimodal.
Who makes Gemini 2.5 Flash (batch)?
Gemini 2.5 Flash (batch) is developed and served by Google.