Mistral Small 4
Efficient, cheap-to-run open model (Apache 2.0).
Modality
Text
Context
Efficient context
Pricing (per 1M)
Open / self-host
What Mistral Small 4 is for
Mistral Small 4 is the efficient tier — built to do useful work at low cost and low latency, and to run on modest hardware where deployed directly.
Mistral's design philosophy is most visible at this size: extracting more from fewer parameters, which makes their small models unusually competitive against larger rivals.
Where it does well
Price and speed without the sharp capability cliff that usually comes with them. For classification, extraction, routing and short-form generation it is genuinely capable rather than merely adequate.
Multilingual handling is a real advantage at this size — most small models are heavily English-weighted and degrade noticeably in other languages.
Where it falls short
It remains a small model. Complex reasoning, long context and nuanced writing are all beyond it, and it will attempt them confidently rather than declining.
Prompts need to be more specific than a larger tier requires, and output should be validated in production.
Should you use it?
A strong default for high-volume, well-defined tasks, particularly multilingual ones. Pair it with Large 3 behind a routing decision rather than using either alone — the cheap model handling the easy majority is the pattern that keeps quality up and cost down.
Capabilities
Pros
- Very efficient
- Cheap to self-host
- Open license
Cons
- Lightweight on complex reasoning
Best for
Try a prompt for Mistral Small 4
Classify this email: support, sales, or spam. One word answer.
Other Mistral models
Pricing indicative, per 1M tokens, reviewed July 2026.
Visit official site