LowRouter
Mistral AI's launch blog "Introducing Mistral 3" (December 2, 2025) reveals the Ministral 3 family of three dense models at 14B, 8B, and 3B sizes, each offered in Base, Instruct, and Reasoning variants for nine total releases under Apache 2.0. The post frames Ministral models as offering the best performance-to-cost ratio in their category versus the headline Mistral Large 3. All Mistral 3 models were trained on NVIDIA Hopper GPUs, and Mistral partnered with NVIDIA, vLLM, and Red Hat to deliver NVFP4 checkpoints, optimized inference on Blackwell NVL72 systems, and deployment through TensorRT-LLM, vLLM, SGLang, llama.cpp, and Ollama. The blog notes a reasoning version of Mistral Large 3 is coming soon, though it does not isolate 14B-specific benchmark numbers.
