Claude Fable 5$22.000/MClaude Opus 5$11.000/MClaude Opus 4.8$11.000/MClaude Opus 4.7$11.000/MClaude Opus 4.6$11.000/MClaude Opus 4.5$33.000/MClaude Sonnet 3.7$6.600/MClaude Opus 3$33.000/MClaude 2.1$12.800/MClaude 2$12.800/MGPT-5.6 Sol$12.500/MGPT-5.6 Terra$5.000/MGPT-5.5$12.500/MGPT-5.2$5.425/MGPT-5.2-Codex$5.425/MGPT-5$3.875/MGPT-4.5$97.500/MGPT-4 Turbo Preview$16.000/MGPT-4$39.000/MGPT-4-32k$78.000/Mo3$19.000/Mo3-mini$2.090/Mo4-mini$2.090/Mo1$28.500/Mo1-mini$5.700/Mo1-preview$28.500/MGemini 3.5 Pro$5.000/MGemini 3.1 Pro$5.000/MGemini 3 Pro$5.000/MGemini 2.5 Pro$3.875/MClaude Fable 5$22.000/MClaude Opus 5$11.000/MClaude Opus 4.8$11.000/MClaude Opus 4.7$11.000/MClaude Opus 4.6$11.000/MClaude Opus 4.5$33.000/MClaude Sonnet 3.7$6.600/MClaude Opus 3$33.000/MClaude 2.1$12.800/MClaude 2$12.800/MGPT-5.6 Sol$12.500/MGPT-5.6 Terra$5.000/MGPT-5.5$12.500/MGPT-5.2$5.425/MGPT-5.2-Codex$5.425/MGPT-5$3.875/MGPT-4.5$97.500/MGPT-4 Turbo Preview$16.000/MGPT-4$39.000/MGPT-4-32k$78.000/Mo3$19.000/Mo3-mini$2.090/Mo4-mini$2.090/Mo1$28.500/Mo1-mini$5.700/Mo1-preview$28.500/MGemini 3.5 Pro$5.000/MGemini 3.1 Pro$5.000/MGemini 3 Pro$5.000/MGemini 2.5 Pro$3.875/M
Meta
Meta
Frontier

Llama 4 Maverick

Frontier

The flagship Llama 4, an MoE-architecture model designed for cheap, high-throughput inference across the open-weights ecosystem.

Llama 4 Maverick is a frontier AI model from Meta. It costs $0.200 per million input tokens and $0.600 per million output tokens (blended $0.320/M), with a 128,000-token context window.

INPUT
$0.200/M
per million input tokens
OUTPUT
$0.600/M
per million output tokens
BLENDED 70/30
$0.320/M
unchanged since 3 May
CONTEXT
128,000
tokens
What it is good at
  • MoE for cheap inference
  • Open weights
  • Wide hosted availability
  • Multimodal
Typical use cases
  • Self-hosted general chat
  • Multi-cloud deployments
  • Fine-tune base

Benchmarks

vs. best public score
MMLU85%
Multitask academic knowledge across 57 subjects.
Graduate-level science questions, "Google-proof".
MATH84%
High-school competition math problems.
Python function synthesis from docstrings.
Real GitHub issues solved end-to-end.
LMArena Elo1310 Elo
Crowd-sourced head-to-head preference Elo rating.
Hand-curated from each provider's published reports and public leaderboards. Methodology varies across sources, treat as directional rather than authoritative.

More from Meta

See all 20

Frequently asked questions

How much does Llama 4 Maverick cost?

Llama 4 Maverick costs $0.200 per million input tokens and $0.600 per million output tokens, for a blended reference rate of $0.320 per million tokens.

What is Llama 4 Maverick's context window?

Llama 4 Maverick supports up to 128,000 tokens of context in a single request.

What is Llama 4 Maverick best for?

Llama 4 Maverick is well suited to MoE for cheap inference, Open weights and Wide hosted availability.

Who makes Llama 4 Maverick?

Llama 4 Maverick is developed and served by Meta.

Compared with

Terms used on this page