Claude Fable 5$22.000/MClaude Opus 5$11.000/MClaude Opus 4.8$11.000/MClaude Opus 4.7$11.000/MClaude Opus 4.6$11.000/MClaude Opus 4.5$33.000/MClaude Sonnet 3.7$6.600/MClaude Opus 3$33.000/MClaude 2.1$12.800/MClaude 2$12.800/MGPT-5.6 Sol$12.500/MGPT-5.6 Terra$5.000/MGPT-5.5$12.500/MGPT-5.2$5.425/MGPT-5.2-Codex$5.425/MGPT-5$3.875/MGPT-4.5$97.500/MGPT-4 Turbo Preview$16.000/MGPT-4$39.000/MGPT-4-32k$78.000/Mo3$19.000/Mo3-mini$2.090/Mo4-mini$2.090/Mo1$28.500/Mo1-mini$5.700/Mo1-preview$28.500/MGemini 3.5 Pro$5.000/MGemini 3.1 Pro$5.000/MGemini 3 Pro$5.000/MGemini 2.5 Pro$3.875/MClaude Fable 5$22.000/MClaude Opus 5$11.000/MClaude Opus 4.8$11.000/MClaude Opus 4.7$11.000/MClaude Opus 4.6$11.000/MClaude Opus 4.5$33.000/MClaude Sonnet 3.7$6.600/MClaude Opus 3$33.000/MClaude 2.1$12.800/MClaude 2$12.800/MGPT-5.6 Sol$12.500/MGPT-5.6 Terra$5.000/MGPT-5.5$12.500/MGPT-5.2$5.425/MGPT-5.2-Codex$5.425/MGPT-5$3.875/MGPT-4.5$97.500/MGPT-4 Turbo Preview$16.000/MGPT-4$39.000/MGPT-4-32k$78.000/Mo3$19.000/Mo3-mini$2.090/Mo4-mini$2.090/Mo1$28.500/Mo1-mini$5.700/Mo1-preview$28.500/MGemini 3.5 Pro$5.000/MGemini 3.1 Pro$5.000/MGemini 3 Pro$5.000/MGemini 2.5 Pro$3.875/M
Groq
Groq
Efficient

Gemma 2 9B (Groq)

Fast Compact

Mid-size Gemma 2, popular open model in the 9B class for on-prem chat and fine-tuning.

Gemma 2 9B (Groq) is a efficient AI model from Groq. It costs $0.200 per million input tokens and $0.200 per million output tokens (blended $0.200/M), with a 8,000-token context window.

Profile inherited from upstream Gemma 2 9B — this is a hosted variant of the same open-weights model.

INPUT
$0.200/M
per million input tokens
OUTPUT
$0.200/M
per million output tokens
BLENDED 70/30
$0.200/M
unchanged since 3 May
CONTEXT
8,000
tokens
What it is good at
  • Single-GPU friendly
  • Open weights
  • Wide ecosystem
Typical use cases
  • On-prem chat
  • Fine-tune base

Benchmarks

vs. best public score
Scores inherited from Gemma 2 9B — this is a hosted variant of the same open-weights model, so the underlying benchmark scores are identical.
MMLU71%
Multitask academic knowledge across 57 subjects.
Graduate-level science questions, "Google-proof".
MATH45%
High-school competition math problems.
Python function synthesis from docstrings.
LMArena Elo1192 Elo
Crowd-sourced head-to-head preference Elo rating.
Hand-curated from each provider's published reports and public leaderboards. Methodology varies across sources, treat as directional rather than authoritative.

More from Groq

See all 11

Frequently asked questions

How much does Gemma 2 9B (Groq) cost?

Gemma 2 9B (Groq) costs $0.200 per million input tokens and $0.200 per million output tokens, for a blended reference rate of $0.200 per million tokens.

What is Gemma 2 9B (Groq)'s context window?

Gemma 2 9B (Groq) supports up to 8,000 tokens of context in a single request.

What is Gemma 2 9B (Groq) best for?

Gemma 2 9B (Groq) is well suited to Single-GPU friendly, Open weights and Wide ecosystem.

Who makes Gemma 2 9B (Groq)?

Gemma 2 9B (Groq) is developed and served by Groq. It was released in Jun 2024.

Terms used on this page