Currently listed through these providers:
Model details
Mistral Large 3 675B Instruct 2512
Mistral Large 3 is built around a granular Mixture-of-Experts architecture pairing a 673B-parameter language backbone with a 2.5B vision encoder, allowing it to process both text and images within a single unified model. Rather than activating all parameters for every token, the model selectively engages only 39B active parameters at inference time, which is what gives the overall 675B design its efficiency edge. This architecture was conceived for real-world deployments: the vision capability lets it analyze documents and visual content alongside text, while native function calling and structured JSON output make it well-suited for agentic pipelines and tool-augmented applications. The multilingual support covers dozens of languages, and the model maintains strong adherence to system prompts, which together make it flexible enough to serve as a daily-driver assistant across diverse user bases.
The model was trained from the ground up using a large cluster of 3000 H200 GPUs, then underwent instruct post-training to sharpen its performance on instruction-following tasks. It ships in FP8 precision for high-throughput serving on modern accelerator hardware, with alternative NVFP4 and BF16 formats available for deployment on H100 or A100 infrastructure. Being released under the Apache 2.0 license, organizations can deploy it on-premises without licensing fees, which lowers the barrier for enterprise adoption in sensitive or regulated environments. The combination of long-context comprehension, open weights, and production-grade reliability positions Mistral Large 3 as a practical foundation for building retrieval-augmented assistants, scientific analysis pipelines, and complex multi-step workflows.
Quick Info
Powered by- Provider
- Nvidia
- Model key
- mistralai/mistral-large-3-675b-instruct-2512
- Release date
- Dec 2, 2025
- Last updated
- Dec 2, 2025
- Knowledge cutoff
- 2025-01
- Input modalities
- Output modalities
- Capabilities
Cost
A provider subscription or plan supersedes token-based pricing for this model.
Limits
- Output tokens
- 262,144 tokens
- Context window
- 262,144 tokens
Latest news about Mistral Large 3 675B Instruct 2512
No articles yet. Fetch the latest news to show it here.