Gemini 2.5 Flash Lite (batch)
Speed-optimised Gemini for high-throughput production. Same multimodal surface as Pro at a fraction of the price.
Gemini 2.5 Flash Lite (batch) is a multimodal AI model from Google. It costs $0.050 per million input tokens and $0.200 per million output tokens (blended $0.095/M), with a 1,048,576-token context window.
Profile inherited from upstream Gemini 2.5 Flash ↗ — this is a hosted variant of the same open-weights model.
- Fast
- 1M context
- Native multimodal
- Very cheap
- Production chat
- Bulk multimodal pipelines
- Cost-sensitive RAG
Benchmarks
More from Google
See all 24 →Frequently asked questions
How much does Gemini 2.5 Flash Lite (batch) cost?
Gemini 2.5 Flash Lite (batch) costs $0.050 per million input tokens and $0.200 per million output tokens, for a blended reference rate of $0.095 per million tokens.
What is Gemini 2.5 Flash Lite (batch)'s context window?
Gemini 2.5 Flash Lite (batch) supports up to 1,048,576 tokens of context in a single request.
What is Gemini 2.5 Flash Lite (batch) best for?
Gemini 2.5 Flash Lite (batch) is well suited to Fast, 1M context and Native multimodal.
Who makes Gemini 2.5 Flash Lite (batch)?
Gemini 2.5 Flash Lite (batch) is developed and served by Google.