Currently listed through these providers:
Model details
Mistral Medium 3
Mistral Medium 3 is built as a frontier-class multimodal language model aimed squarely at enterprise workloads where high accuracy, manageable infrastructure, and cost control all matter at once. The 25.05 release pairs vision understanding with an extended 128k context length, letting a single model read and reason over long documents, slide decks, and mixed visual-text inputs rather than splitting those jobs across separate systems. Mistral positions it as a new "medium" tier that intentionally closes the gap with much larger frontier models while keeping the compute footprint light enough to run on a modest hardware stack such as 4xH100 GPUs, supporting both cloud and on-premises or in-VPC deployment. Its design intent shows up in the capability mix: tool calling, attachments, and temperature control are first-class, so it can drop into agentic pipelines and coding assistants as the central reasoning engine without needing bespoke scaffolding around it.
The model is framed by its creators as a balance point between open-weights experimentation and closed enterprise services, carrying forward the lineage of Mistral's earlier open releases while adding the post-training and integration polish typically expected of an enterprise product. Independent coverage notes that it performs at or above 90% of larger frontier competitors on a broad range of benchmarks while being an order of magnitude cheaper to run, and that it surpasses leading open and enterprise peers on professional tasks such as coding and multimodal understanding. Its qualitative strengths line up with practical enterprise needs: it can be deployed in private or hybrid environments, customized through post-training, and woven into existing toolchains, which makes it a natural fit for organizations that want a versatile long-context model without paying frontier prices. For teams building agents, document understanding systems, or coding copilots in 2025 and beyond, Mistral Medium 3 represents a forward-looking middle ground where multimodal understanding, long context, and efficient deployment converge in a single model.
Quick Info
Powered by- Provider
- Merge Gateway
- Model key
- mistral/mistral-medium-2505
- Release date
- May 7, 2025
- Last updated
- May 7, 2025
- Knowledge cutoff
- 2025-05
- Input modalities
- Output modalities
- Capabilities
Cost
- Input token cost
- $0.40
- Output token cost
- $2.00
Limits
- Output tokens
- 128,000 tokens
- Context window
- 128,000 tokens
Transparent token rates
Compare mistral-medium pricing
Rates are shown per one million tokens. Combined means one million input plus one million output tokens.
Latest news about Mistral Medium 3
Videos about Mistral Medium 3
More models around Mistral Medium 3
This exact model name is also listed by 7 other providers.