LowRouter
Mistral AI announced Mistral Small 4 on March 16, 2026, unifying the capabilities of Magistral reasoning, Pixtral multimodal, and Devstral agentic coding into a single hybrid model. It is a 119B-parameter Mixture-of-Experts model with 128 experts and 4 active per token, supporting a 256k context window, configurable reasoning effort, and native text-and-image inputs. The model is released under the Apache 2.0 license, and Mistral joined the NVIDIA Nemotron Coalition as a founding member alongside the launch. Performance claims versus Mistral Small 3 include a 40% reduction in end-to-end completion time under a latency-optimized setup and 3x more requests per second in a throughput-optimized setup. Mistral Small 4 activates roughly 6B parameters per token (8B including embeddings), letting users toggle between fast instruct responses and deeper reasoning-intensive outputs on demand. The release positions Small 4 as a single versatile model for chat, coding, agentic tasks, and complex reasoning.
