Gemma 4 31B
Gemma 4 31B is a general-purpose chat-tuned model. Available via Google. 262,144-token context. Budget pricing.
Gemma 4 31B is a multimodal AI model from Google. It costs $0.090 per million input tokens and $0.340 per million output tokens (blended $0.165/M), with a 262,144-token context window.
+11.8%over 91 days · hover to read
- Solid general chat performance
- Reasonable instruction following
- Multimodal: handles images alongside text
- General chat
- Drafting and rewriting
- Personal productivity
Benchmarks
More from Google
See all 24 →Price history
4 recorded changes| Date | Rate | Was | Now | Change |
|---|---|---|---|---|
| 25 Jul 2026 | Input | $0.120 | $0.140 | +17% |
| 25 Jul 2026 | Output | $0.370 | $0.400 | +8.1% |
| 12 May 2026 | Input | $0.130 | $0.120 | -7.7% |
| 12 May 2026 | Output | $0.380 | $0.370 | -2.6% |
Rates per million tokens, as published by the provider or its router listing. Tracking runs since 3 May; anything earlier is not on record. See every tracked price change.
Frequently asked questions
How much does Gemma 4 31B cost?
Gemma 4 31B costs $0.090 per million input tokens and $0.340 per million output tokens, for a blended reference rate of $0.165 per million tokens.
What is Gemma 4 31B's context window?
Gemma 4 31B supports up to 262,144 tokens of context in a single request.
What is Gemma 4 31B best for?
Gemma 4 31B is well suited to Solid general chat performance, Reasonable instruction following and Multimodal: handles images alongside text.
Who makes Gemma 4 31B?
Gemma 4 31B is developed and served by Google.