Currently listed through these providers:
Model details
MiniMax-M2.1
MiniMax-M2.1 sits within the MiniMax family as a coding-focused model designed for demanding, multi-language programming work. Documentation references describe it under a legacy heading while newer members like M2.7 and the multimodal M3 have since taken the spotlight, and a Hacker News discussion links to a MiniMax news post introducing it as "Built for Real-World Complex Tasks, Multi-Language Programming." That positioning suggests the model was framed less as a general conversational assistant and more as a workhorse for developers juggling polyglot codebases, refactors, and agent-style workflows where tool calling and code reasoning matter.
The public-facing description highlighted a sparse architecture with around 230 billion total parameters but only about 10 billion activated per inference, a design trade-off that targets throughput without giving up the capacity to handle long, real-world engineering sessions. Practitioner commentary in the Hacker News thread placed its coding ability in the neighborhood of contemporary Sonnet-class models, while praising its pricing aggressiveness enough to justify running multiple parallel passes. Fit-wise, the model is well suited to teams that need open weights for self-hosting, want a long context window for navigating whole repositories, and care more about reliable code generation and refactoring than cutting-edge general reasoning.
Quick Info
Powered by- Provider
- MiniMax (minimax.io)
- Model key
- MiniMax-M2.1
- Release date
- Dec 23, 2025
- Last updated
- Dec 23, 2025
- Input modalities
- Output modalities
- Capabilities
Cost
- Input token cost
- $0.30
- Output token cost
- $1.20
Limits
- Output tokens
- 131,072 tokens
- Context window
- 204,800 tokens