Gemma 3n 4B
text->text
Gemma 3n 4B is a general-purpose chat-tuned model. Available via Google. 32,768-token context. Budget pricing.
Gemma 3n 4B is a efficient AI model from Google. It costs $0.060 per million input tokens and $0.120 per million output tokens (blended $0.078/M), with a 32,768-token context window.
INPUT
$0.060/M
per million input tokens
OUTPUT
$0.120/M
per million output tokens
CONTEXT
32,768
tokens
What it is good at
- Solid general chat performance
- Reasonable instruction following
Typical use cases
- General chat
- Drafting and rewriting
- Personal productivity
Benchmarks
No published benchmark scores tracked for this model yet. Frontier and reasoning models from major providers have scores; smaller models and inference-host variants typically inherit the underlying open-weights score.
More from Google
See all 24 →Gemini 3.5 Pro
Frontier · 1,000,000 ctx
in $2.000/Mout $12.000/M
Gemini 3.6 Flash
Balanced · 1,050,000 ctx
in $1.500/Mout $7.500/M
Gemini 3.5 Flash
Balanced · 1,000,000 ctx
in $1.500/Mout $9.000/M
Gemini 3.1 Pro
Frontier · 1,000,000 ctx
in $2.000/Mout $12.000/M
Gemini 3 Pro
Frontier · 1,000,000 ctx
in $2.000/Mout $12.000/M
Gemini 3 Flash
Efficient · 1,000,000 ctx
in $0.500/Mout $3.000/M
Frequently asked questions
How much does Gemma 3n 4B cost?
Gemma 3n 4B costs $0.060 per million input tokens and $0.120 per million output tokens, for a blended reference rate of $0.078 per million tokens.
What is Gemma 3n 4B's context window?
Gemma 3n 4B supports up to 32,768 tokens of context in a single request.
What is Gemma 3n 4B best for?
Gemma 3n 4B is well suited to Solid general chat performance, Reasonable instruction following and General chat.
Who makes Gemma 3n 4B?
Gemma 3n 4B is developed and served by Google.