Currently listed through these providers:
Model details
MiniMax-M3
MiniMax-M3 is presented by its publisher as a frontier coding model that combines a one-million-token context window with native multimodality in a single architecture. Independent reference documentation describes it as a vision-language model built on a Mixture-of-Experts design, allowing it to ingest text, images, and video while emitting text outputs. That combination positions the system as a long-context reasoning engine rather than a narrow chat model, aimed at workflows that need to hold large codebases, design assets, or lengthy video transcripts in working memory at the same time.
The practical emphasis for MiniMax-M3 falls on long-horizon software engineering, agentic task execution, and creative production work. The reference description highlights long-form video understanding, sustained coding sessions across many files, and design or creative pipelines as core scenarios, suggesting the MoE routing is tuned for sustained reasoning and tool use rather than single-turn answers. Distribution through NVIDIA NIM under a non-commercial license and links to a public model card indicate that weights and serving recipes are openly accessible to developers who want to self-host or experiment, making it a flexible foundation for teams building assistants that blend code generation, visual analysis, and extended context reasoning.
Quick Info
Powered by- Provider
- Volcengine Ark Coding Plan
- Model key
- minimax-m3
- Release date
- Jun 1, 2026
- Last updated
- Jun 1, 2026
- Input modalities
- Output modalities
- Capabilities
Cost
A provider subscription or plan supersedes token-based pricing for this model.
Limits
- Output tokens
- 512,000 tokens
- Context window
- 1,048,576 tokens
Latest news about MiniMax-M3
Videos about MiniMax-M3
More models around MiniMax-M3
This exact model name is also listed by 39 other providers.