Mistral Small 3.2
Efficient
Mistral's 24B small model: open-weights, fast, and strong on instruction following for its size.
Mistral Small 3.2 is a efficient AI model from Mistral. It costs $0.100 per million input tokens and $0.300 per million output tokens (blended $0.160/M), with a 128,000-token context window.
INPUT
$0.100/M
per million input tokens
OUTPUT
$0.300/M
per million output tokens
CONTEXT
128,000
tokens
What it is good at
- Open weights (Apache 2.0)
- Single-GPU friendly
- 128K context
Typical use cases
- Self-hosted production chat
- Cheap European-hosted inference
Benchmarks
vs. best public score
Hand-curated from each provider's published reports and public leaderboards. Methodology varies across sources, treat as directional rather than authoritative.
More from Mistral
See all 15 →Mistral Large 3
Frontier · 256,000 ctx
in $0.500/Mout $1.500/M
Mistral Medium 3.5
Balanced · 256,000 ctx
in $1.500/Mout $7.500/M
Mistral Small 4
Efficient · 128,000 ctx
in $0.100/Mout $0.300/M
Mistral Large 2
Frontier · 128,000 ctx
in $2.000/Mout $6.000/M
Pixtral Large
Vision · 128,000 ctx
in $2.000/Mout $6.000/M
Pixtral 12B
Compact Vision · 128,000 ctx
in $0.150/Mout $0.150/M
Frequently asked questions
How much does Mistral Small 3.2 cost?
Mistral Small 3.2 costs $0.100 per million input tokens and $0.300 per million output tokens, for a blended reference rate of $0.160 per million tokens.
What is Mistral Small 3.2's context window?
Mistral Small 3.2 supports up to 128,000 tokens of context in a single request.
What is Mistral Small 3.2 best for?
Mistral Small 3.2 is well suited to Open weights (Apache 2.0), Single-GPU friendly and 128K context.
Who makes Mistral Small 3.2?
Mistral Small 3.2 is developed and served by Mistral.