Mistral
On December 2, 2025, Mistral announced the Mistral 3 generation, which includes three dense Ministral 3 small models (14B, 8B, and 3B) alongside Mistral Large 3, all released under Apache 2.0. Mistral framed the Ministral 3 family as offering the best performance-to-cost ratio in its category for developers. Mistral partnered with NVIDIA, vLLM, and Red Hat to optimize the new models, training them on NVIDIA Hopper GPUs and shipping NVFP4 checkpoints via llm-compressor for efficient inference on Blackwell NVL72 systems and 8×A100 or 8×H100 vLLM nodes. The launch emphasizes open-sourcing compressed formats to enable distributed intelligence across the community.
