Mistral Small 3
Mistral's 24B small model: open-weights, fast, and strong on instruction following for its size.
Mistral Small 3 is a efficient AI model from Mistral. It costs $0.050 per million input tokens and $0.080 per million output tokens (blended $0.059/M), with a 32,768-token context window.
Profile inherited from upstream Mistral Small 3.2 ↗ — this is a hosted variant of the same open-weights model.
- Open weights (Apache 2.0)
- Single-GPU friendly
- 128K context
- Self-hosted production chat
- Cheap European-hosted inference
Benchmarks
More from Mistral
See all 15 →Frequently asked questions
How much does Mistral Small 3 cost?
Mistral Small 3 costs $0.050 per million input tokens and $0.080 per million output tokens, for a blended reference rate of $0.059 per million tokens.
What is Mistral Small 3's context window?
Mistral Small 3 supports up to 32,768 tokens of context in a single request.
What is Mistral Small 3 best for?
Mistral Small 3 is well suited to Open weights (Apache 2.0), Single-GPU friendly and 128K context.
Who makes Mistral Small 3?
Mistral Small 3 is developed and served by Mistral.