Sulat.com
AI models
Mistral logo

Model details

Mistral Small 4

Mistral Small 4 is designed as a unified hybrid system that merges the distinct strengths of instruction-following, reasoning, and coding-focused model families. Built on a Mixture-of-Experts architecture, it utilizes 128 experts with 4 active at any given time, resulting in a total of 119 billion parameters while maintaining efficiency with only 6.5 billion parameters activated per token. This design intent allows the model to function as a flexible, general-purpose tool capable of handling both text and image inputs, making it well-suited for complex tasks that require a blend of visual analysis and logical processing.

The model lineage emphasizes adaptability through configurable reasoning effort, allowing users to toggle between rapid responses and deeper, compute-intensive analysis. To support diverse deployment needs, the architecture is optimized for performance, offering significant improvements in latency and throughput compared to previous generations. Developers can further refine efficiency through techniques like speculative decoding with specialized eagle heads or by utilizing 4-bit float precision quantization. These advancements position the model as a robust choice for agentic workflows and multilingual applications that demand high performance and architectural flexibility.

Mistralmistral-small-2603mistral-small

Quick Info

Powered by
Provider
Mistral
Model key
mistral-small-2603
Release date
Mar 16, 2026
Last updated
Mar 16, 2026
Knowledge cutoff
2025-06
Input modalities
Output modalities
Capabilities

Cost

Input token cost
$0.15
Output token cost
$0.60

Limits

Output tokens
256,000 tokens
Context window
256,000 tokens

Latest news about Mistral Small 4

Mistral

Official sourceAnnouncement

A new standard for multimodal, reasoning-optimized models. Mistral Small 4 is a hybrid model optimized for general chat, coding, agentic tasks, and complex ...

Videos about Mistral Small 4

More models around Mistral Small 4