Gemma 2 2B
Nano
Smallest Gemma 2, runs on a phone or laptop. Strong tiny-model baseline.
Gemma 2 2B is a efficient AI model from Google. It costs $0.020 per million input tokens and $0.020 per million output tokens (blended $0.020/M), with a 8,000-token context window.
INPUT
$0.020/M
per million input tokens
OUTPUT
$0.020/M
per million output tokens
CONTEXT
8,000
tokens
What it is good at
- Edge / mobile capable
- Open weights
- 2B params
Typical use cases
- On-device assistants
- Edge inference
Benchmarks
vs. best public score
Hand-curated from each provider's published reports and public leaderboards. Methodology varies across sources, treat as directional rather than authoritative.
More from Google
See all 24 →Gemini 3.5 Pro
Frontier · 1,000,000 ctx
in $2.000/Mout $12.000/M
Gemini 3.6 Flash
Balanced · 1,050,000 ctx
in $1.500/Mout $7.500/M
Gemini 3.5 Flash
Balanced · 1,000,000 ctx
in $1.500/Mout $9.000/M
Gemini 3.1 Pro
Frontier · 1,000,000 ctx
in $2.000/Mout $12.000/M
Gemini 3 Pro
Frontier · 1,000,000 ctx
in $2.000/Mout $12.000/M
Gemini 3 Flash
Efficient · 1,000,000 ctx
in $0.500/Mout $3.000/M
Frequently asked questions
How much does Gemma 2 2B cost?
Gemma 2 2B costs $0.020 per million input tokens and $0.020 per million output tokens, for a blended reference rate of $0.020 per million tokens.
What is Gemma 2 2B's context window?
Gemma 2 2B supports up to 8,000 tokens of context in a single request.
What is Gemma 2 2B best for?
Gemma 2 2B is well suited to Edge / mobile capable, Open weights and 2B params.
Who makes Gemma 2 2B?
Gemma 2 2B is developed and served by Google. It was released in Jul 2024.