Gemini 2.5 Flash Lite
Speed-optimised Gemini for high-throughput production. Same multimodal surface as Pro at a fraction of the price.
Gemini 2.5 Flash Lite is a multimodal AI model from Google. It costs $0.100 per million input tokens and $0.400 per million output tokens (blended $0.190/M), with a 1,048,576-token context window.
Profile inherited from upstream Gemini 2.5 Flash ↗ — this is a hosted variant of the same open-weights model.
- Fast
- 1M context
- Native multimodal
- Very cheap
- Production chat
- Bulk multimodal pipelines
- Cost-sensitive RAG
Benchmarks
More from Google
See all 24 →Frequently asked questions
How much does Gemini 2.5 Flash Lite cost?
Gemini 2.5 Flash Lite costs $0.100 per million input tokens and $0.400 per million output tokens, for a blended reference rate of $0.190 per million tokens.
What is Gemini 2.5 Flash Lite's context window?
Gemini 2.5 Flash Lite supports up to 1,048,576 tokens of context in a single request.
What is Gemini 2.5 Flash Lite best for?
Gemini 2.5 Flash Lite is well suited to Fast, 1M context and Native multimodal.
Who makes Gemini 2.5 Flash Lite?
Gemini 2.5 Flash Lite is developed and served by Google.