Model details
mistral-large-2512
Mistral Large 3 is a general-purpose multimodal model built on a granular mixture-of-experts architecture. It utilizes 675B total parameters, with 41B active parameters engaged during inference to balance performance and efficiency. Designed to handle diverse and demanding workloads, the model supports a 256k token context window and native vision capabilities, making it well-suited for complex reasoning, document analysis, and large-scale data processing.
The model was trained from the ground up using a cluster of 3000 H200 GPUs, reflecting a significant investment in computational scale. Following its initial training, the model underwent instruction-based post-training to refine its performance for chat, agentic workflows, and specialized instruction-following tasks. Available in various precision formats, including FP8 and BF16, it is engineered for production-grade reliability, allowing for deployment in both cloud environments and on-premises infrastructure for enterprise-level applications.
Quick Info
Powered by- Provider
- 302.AI
- Model key
- mistral-large-2512
- Release date
- Dec 16, 2025
- Last updated
- Dec 16, 2025
- Knowledge cutoff
- 2024-12
- Input modalities
- Output modalities
- Capabilities
Cost
- Input token cost
- $1.10
- Output token cost
- $3.30
Limits
- Output tokens
- 262,144 tokens
- Context window
- 128,000 tokens