LowRouter
Mistral AI announced Mistral 3 on December 2, 2025, headlined by Mistral Large 3, a sparse mixture-of-experts model with 41B active and 675B total parameters, released with base and instruction-tuned weights under Apache 2.0. It is Mistral's first MoE since the Mixtral series, trained from scratch on 3000 NVIDIA H200 GPUs, and debuted at #2 on LMArena among open-source non-reasoning models. The launch confirmed image understanding and best-in-class multilingual conversation performance outside English and Chinese, with a reasoning version to follow. Mistral partnered with NVIDIA, vLLM, and Red Hat to ship an NVFP4 checkpoint built with llm-compressor, enabling efficient runs on Blackwell NVL72 systems or a single 8×A100 / 8×H100 node. All Mistral 3 models were trained on NVIDIA Hopper GPUs.
