Currently listed through these providers:
Model details
Mistral Small 4
Mistral Small 4 is presented in Mistral's official model overview as a featured release positioned in the Small tier, sitting alongside Mistral Medium 3.5 and the Voxtral transcription models. According to the documentation listing, it is described as a hybrid model intended to unify instruction-following, reasoning, and coding behaviour within a single efficient checkpoint, suggesting a design goal of replacing separate instruct and reasoning variants with one configurable system. The corresponding NVIDIA NGC catalogue entry, named mistral-small-4-119b-2603 and published under the mistralai namespace on NIM, indicates a packaging around a roughly 119-billion-parameter variant that organisations can deploy on NVIDIA infrastructure, giving enterprise users a standardised runtime path for the same model surfaced through Mistral's own APIs.
In practical terms, Mistral Small 4 is best understood as a mid-sized, versatile workhorse rather than a frontier-class model. The hybrid framing on the official overview points to flexible chat-style usage that can lean on its coding strengths for developer assistance and tool-driven workflows, while still handling general reasoning and instruction tasks that would otherwise require a larger model. Because the model is also distributed through an NVIDIA NIM container, it fits well into teams that already standardise on NVIDIA hardware and want a reproducible deployment artefact, while the underlying Mistral lineage and the Small-tier positioning suggest a favourable balance of quality, latency, and operating cost for everyday assistant, automation, and coding workloads.
Quick Info
Powered by- Provider
- EmpirioLabs AI
- Model key
- mistral-small-4
- Release date
- Mar 16, 2026
- Last updated
- Mar 16, 2026
- Knowledge cutoff
- 2025-06
- Input modalities
- Output modalities
- Capabilities
Cost
- Input token cost
- $0.15
- Output token cost
- $0.60
Limits
- Output tokens
- 65,536 tokens
- Context window
- 256,000 tokens
Transparent token rates
Compare mistral-small pricing
Rates are shown per one million tokens. Combined means one million input plus one million output tokens.
Latest news about Mistral Small 4
Videos about Mistral Small 4
More models around Mistral Small 4
This exact model name is also listed by 11 other providers.