Llama 4 Maverick
The flagship Llama 4, an MoE-architecture model designed for cheap, high-throughput inference across the open-weights ecosystem.
Llama 4 Maverick is a frontier AI model from Meta. It costs $0.200 per million input tokens and $0.600 per million output tokens (blended $0.320/M), with a 128,000-token context window.
0.0%over 27 days · hover to read
- MoE for cheap inference
- Open weights
- Wide hosted availability
- Multimodal
- Self-hosted general chat
- Multi-cloud deployments
- Fine-tune base
Benchmarks
More from Meta
See all 20 →Price history
2 recorded changes| Date | Rate | Was | Now | Change |
|---|---|---|---|---|
| 15 Sept 2026 | Output | $0.600 | $0.652 | +8.8% |
| 15 Sept 2026 | Input | $0.150 | $0.188 | +25% |
Rates per million tokens, as published by the provider or its router listing. Tracking runs since 3 May; anything earlier is not on record. See every tracked price change.
Frequently asked questions
How much does Llama 4 Maverick cost?
Llama 4 Maverick costs $0.200 per million input tokens and $0.600 per million output tokens, for a blended reference rate of $0.320 per million tokens.
What is Llama 4 Maverick's context window?
Llama 4 Maverick supports up to 128,000 tokens of context in a single request.
What is Llama 4 Maverick best for?
Llama 4 Maverick is well suited to MoE for cheap inference, Open weights and Wide hosted availability.
Who makes Llama 4 Maverick?
Llama 4 Maverick is developed and served by Meta.