Mistral Small 4 represents a major release in the Mistral Small family, designed as a single hybrid system that combines instruction following, extended reasoning, and coding ability. Built on a Mixture-of-Experts-style architecture with 119 billion total parameters and 6.5 billion active per inference, the model is engineered to handle demanding analytical workloads without requiring users to switch between separate models for different tasks. Released under the Apache 2.0 license as version v26.03 on March 16, 2026, it is open-weight, making it accessible for self-hosting, fine-tuning, and enterprise customization while still being available through managed API endpoints.
In practical terms, Mistral Small 4 is best suited for complex problem solving, mathematical reasoning, planning, and multi-step analysis where chain-of-thought generation adds clear value. The model accepts both text and image inputs, enabling multimodal workflows such as document or diagram understanding paired with structured text reasoning. Its large context window supports long-form analysis, code repositories, and agentic pipelines with function calling and structured output. Trade-offs to consider are higher inference latency and elevated compute cost relative to smaller Mistral variants, so it fits best where reasoning depth matters more than raw response speed, such as research assistance, code generation on hard problems, and orchestrated agent workflows.