Claude Fable 5$22.000/MClaude Opus 5$11.000/MClaude Opus 4.8$11.000/MClaude Opus 4.7$11.000/MClaude Opus 4.6$11.000/MClaude Opus 4.5$33.000/MClaude Sonnet 3.7$6.600/MClaude Opus 3$33.000/MClaude 2.1$12.800/MClaude 2$12.800/MGPT-5.6 Sol$12.500/MGPT-5.6 Terra$5.000/MGPT-5.5$12.500/MGPT-5.2$5.425/MGPT-5.2-Codex$5.425/MGPT-5$3.875/MGPT-4.5$97.500/MGPT-4 Turbo Preview$16.000/MGPT-4$39.000/MGPT-4-32k$78.000/Mo3$19.000/Mo3-mini$2.090/Mo4-mini$2.090/Mo1$28.500/Mo1-mini$5.700/Mo1-preview$28.500/MGemini 3.5 Pro$5.000/MGemini 3.1 Pro$5.000/MGemini 3 Pro$5.000/MGemini 2.5 Pro$3.875/MClaude Fable 5$22.000/MClaude Opus 5$11.000/MClaude Opus 4.8$11.000/MClaude Opus 4.7$11.000/MClaude Opus 4.6$11.000/MClaude Opus 4.5$33.000/MClaude Sonnet 3.7$6.600/MClaude Opus 3$33.000/MClaude 2.1$12.800/MClaude 2$12.800/MGPT-5.6 Sol$12.500/MGPT-5.6 Terra$5.000/MGPT-5.5$12.500/MGPT-5.2$5.425/MGPT-5.2-Codex$5.425/MGPT-5$3.875/MGPT-4.5$97.500/MGPT-4 Turbo Preview$16.000/MGPT-4$39.000/MGPT-4-32k$78.000/Mo3$19.000/Mo3-mini$2.090/Mo4-mini$2.090/Mo1$28.500/Mo1-mini$5.700/Mo1-preview$28.500/MGemini 3.5 Pro$5.000/MGemini 3.1 Pro$5.000/MGemini 3 Pro$5.000/MGemini 2.5 Pro$3.875/M
Google
Google
MultimodalLIVE INDEX

Gemini 2.5 Flash Lite (batch)

text+image+file+audio+video->text

Speed-optimised Gemini for high-throughput production. Same multimodal surface as Pro at a fraction of the price.

Gemini 2.5 Flash Lite (batch) is a multimodal AI model from Google. It costs $0.050 per million input tokens and $0.200 per million output tokens (blended $0.095/M), with a 1,048,576-token context window.

Profile inherited from upstream Gemini 2.5 Flash — this is a hosted variant of the same open-weights model.

INPUT
$0.050/M
per million input tokens
OUTPUT
$0.200/M
per million output tokens
CONTEXT
1,048,576
tokens
What it is good at
  • Fast
  • 1M context
  • Native multimodal
  • Very cheap
Typical use cases
  • Production chat
  • Bulk multimodal pipelines
  • Cost-sensitive RAG

Benchmarks

vs. best public score
Scores inherited from Gemini 2.5 Flash — this is a hosted variant of the same open-weights model, so the underlying benchmark scores are identical.
MMLU84%
Multitask academic knowledge across 57 subjects.
Graduate-level science questions, "Google-proof".
MATH87%
High-school competition math problems.
Python function synthesis from docstrings.
Real GitHub issues solved end-to-end.
LMArena Elo1325 Elo
Crowd-sourced head-to-head preference Elo rating.
Hand-curated from each provider's published reports and public leaderboards. Methodology varies across sources, treat as directional rather than authoritative.

More from Google

See all 24

Frequently asked questions

How much does Gemini 2.5 Flash Lite (batch) cost?

Gemini 2.5 Flash Lite (batch) costs $0.050 per million input tokens and $0.200 per million output tokens, for a blended reference rate of $0.095 per million tokens.

What is Gemini 2.5 Flash Lite (batch)'s context window?

Gemini 2.5 Flash Lite (batch) supports up to 1,048,576 tokens of context in a single request.

What is Gemini 2.5 Flash Lite (batch) best for?

Gemini 2.5 Flash Lite (batch) is well suited to Fast, 1M context and Native multimodal.

Who makes Gemini 2.5 Flash Lite (batch)?

Gemini 2.5 Flash Lite (batch) is developed and served by Google.

Terms used on this page