Mistral
Mistral AI announced Mistral 3 on December 2, 2025, introducing a new generation of models including three dense variants — 14B, 8B, and 3B — alongside Mistral Large 3, a sparse mixture-of-experts model with 41B active and 675B total parameters. All models are released under the Apache 2.0 license. The Ministral models are positioned as offering the best performance-to-cost ratio in their category, targeting efficient deployment for developers and enterprises. All Mistral 3 models, from Large 3 down to Ministral 3, were trained on NVIDIA Hopper GPUs. Mistral partnered with NVIDIA, vLLM, and Red Hat to optimize accessibility, releasing an NVFP4 checkpoint built with llm-compressor that enables efficient inference on Blackwell NVL72 systems and on 8×A100 or 8×H100 nodes. The release also debuts Mistral's first MoE since the Mixtral series, marking a substantial pretraining step forward for the company.
