Mixtral 8x22B
Open MoE
Largest open-weights Mixtral MoE. Cheap-to-serve frontier-ish quality before Llama 3.1 405B and DeepSeek V3 took the open-weights lead.
Mixtral 8x22B is a efficient AI model from Mistral. It costs $0.900 per million input tokens and $0.900 per million output tokens (blended $0.900/M), with a 64,000-token context window.
INPUT
$0.900/M
per million input tokens
OUTPUT
$0.900/M
per million output tokens
CONTEXT
64,000
tokens
What it is good at
- Open-weights MoE
- Cheap inference per active parameter
- 64K context
Typical use cases
- Self-hosted chat at scale
- Fine-tune base
Benchmarks
vs. best public score
Hand-curated from each provider's published reports and public leaderboards. Methodology varies across sources, treat as directional rather than authoritative.
More from Mistral
See all 15 →Mistral Large 3
Frontier · 256,000 ctx
in $0.500/Mout $1.500/M
Mistral Medium 3.5
Balanced · 256,000 ctx
in $1.500/Mout $7.500/M
Mistral Small 4
Efficient · 128,000 ctx
in $0.100/Mout $0.300/M
Mistral Large 2
Frontier · 128,000 ctx
in $2.000/Mout $6.000/M
Pixtral Large
Vision · 128,000 ctx
in $2.000/Mout $6.000/M
Pixtral 12B
Compact Vision · 128,000 ctx
in $0.150/Mout $0.150/M
Frequently asked questions
How much does Mixtral 8x22B cost?
Mixtral 8x22B costs $0.900 per million input tokens and $0.900 per million output tokens, for a blended reference rate of $0.900 per million tokens.
What is Mixtral 8x22B's context window?
Mixtral 8x22B supports up to 64,000 tokens of context in a single request.
What is Mixtral 8x22B best for?
Mixtral 8x22B is well suited to Open-weights MoE, Cheap inference per active parameter and 64K context.
Who makes Mixtral 8x22B?
Mixtral 8x22B is developed and served by Mistral. It was released in Apr 2024.