Claude Fable 5$22.000/MClaude Opus 5$11.000/MClaude Opus 4.8$11.000/MClaude Opus 4.7$11.000/MClaude Opus 4.6$11.000/MClaude Opus 4.5$33.000/MClaude Sonnet 3.7$6.600/MClaude Opus 3$33.000/MClaude 2.1$12.800/MClaude 2$12.800/MGPT-5.6 Sol$12.500/MGPT-5.6 Terra$5.000/MGPT-5.5$12.500/MGPT-5.2$5.425/MGPT-5.2-Codex$5.425/MGPT-5$3.875/MGPT-4.5$97.500/MGPT-4 Turbo Preview$16.000/MGPT-4$39.000/MGPT-4-32k$78.000/Mo3$19.000/Mo3-mini$2.090/Mo4-mini$2.090/Mo1$28.500/Mo1-mini$5.700/Mo1-preview$28.500/MGemini 3.5 Pro$5.000/MGemini 3.1 Pro$5.000/MGemini 3 Pro$5.000/MGemini 2.5 Pro$3.875/MClaude Fable 5$22.000/MClaude Opus 5$11.000/MClaude Opus 4.8$11.000/MClaude Opus 4.7$11.000/MClaude Opus 4.6$11.000/MClaude Opus 4.5$33.000/MClaude Sonnet 3.7$6.600/MClaude Opus 3$33.000/MClaude 2.1$12.800/MClaude 2$12.800/MGPT-5.6 Sol$12.500/MGPT-5.6 Terra$5.000/MGPT-5.5$12.500/MGPT-5.2$5.425/MGPT-5.2-Codex$5.425/MGPT-5$3.875/MGPT-4.5$97.500/MGPT-4 Turbo Preview$16.000/MGPT-4$39.000/MGPT-4-32k$78.000/Mo3$19.000/Mo3-mini$2.090/Mo4-mini$2.090/Mo1$28.500/Mo1-mini$5.700/Mo1-preview$28.500/MGemini 3.5 Pro$5.000/MGemini 3.1 Pro$5.000/MGemini 3 Pro$5.000/MGemini 2.5 Pro$3.875/M
Google
Google
MultimodalLIVE INDEX

Gemini 2.5 Flash (batch)

text+image+file+audio+video->text

Speed-optimised Gemini for high-throughput production. Same multimodal surface as Pro at a fraction of the price.

Gemini 2.5 Flash (batch) is a multimodal AI model from Google. It costs $0.150 per million input tokens and $1.250 per million output tokens (blended $0.480/M), with a 1,048,576-token context window.

Profile inherited from upstream Gemini 2.5 Flash — this is a hosted variant of the same open-weights model.

INPUT
$0.150/M
per million input tokens
OUTPUT
$1.250/M
per million output tokens
CONTEXT
1,048,576
tokens
What it is good at
  • Fast
  • 1M context
  • Native multimodal
  • Very cheap
Typical use cases
  • Production chat
  • Bulk multimodal pipelines
  • Cost-sensitive RAG

Benchmarks

vs. best public score
Scores inherited from Gemini 2.5 Flash — this is a hosted variant of the same open-weights model, so the underlying benchmark scores are identical.
MMLU84%
Multitask academic knowledge across 57 subjects.
Graduate-level science questions, "Google-proof".
MATH87%
High-school competition math problems.
Python function synthesis from docstrings.
Real GitHub issues solved end-to-end.
LMArena Elo1325 Elo
Crowd-sourced head-to-head preference Elo rating.
Hand-curated from each provider's published reports and public leaderboards. Methodology varies across sources, treat as directional rather than authoritative.

More from Google

See all 24

Frequently asked questions

How much does Gemini 2.5 Flash (batch) cost?

Gemini 2.5 Flash (batch) costs $0.150 per million input tokens and $1.250 per million output tokens, for a blended reference rate of $0.480 per million tokens.

What is Gemini 2.5 Flash (batch)'s context window?

Gemini 2.5 Flash (batch) supports up to 1,048,576 tokens of context in a single request.

What is Gemini 2.5 Flash (batch) best for?

Gemini 2.5 Flash (batch) is well suited to Fast, 1M context and Native multimodal.

Who makes Gemini 2.5 Flash (batch)?

Gemini 2.5 Flash (batch) is developed and served by Google.

Terms used on this page