Mistral Small 4
Mistral's 24B small model: open-weights, fast, and strong on instruction following for its size.
Mistral Small 4 is a efficient AI model from Mistral. It costs $0.100 per million input tokens and $0.300 per million output tokens (blended $0.160/M), with a 128,000-token context window.
Profile inherited from upstream Mistral Small 3.2 ↗ — this is a hosted variant of the same open-weights model.
- Open weights (Apache 2.0)
- Single-GPU friendly
- 128K context
- Self-hosted production chat
- Cheap European-hosted inference
Benchmarks
More from Mistral
See all 15 →Frequently asked questions
How much does Mistral Small 4 cost?
Mistral Small 4 costs $0.100 per million input tokens and $0.300 per million output tokens, for a blended reference rate of $0.160 per million tokens.
What is Mistral Small 4's context window?
Mistral Small 4 supports up to 128,000 tokens of context in a single request.
What is Mistral Small 4 best for?
Mistral Small 4 is well suited to Open weights (Apache 2.0), Single-GPU friendly and 128K context.
Who makes Mistral Small 4?
Mistral Small 4 is developed and served by Mistral.