Model details
MiniMax M2.7 highspeed
The MiniMax M2.7 highspeed is built on a sparse Mixture-of-Experts architecture that utilizes 230 billion total parameters while activating only 10 billion per inference. This design intent focuses on achieving high-tier performance comparable to industry-leading models while maintaining significantly lower operational costs. By optimizing the active parameter count, the model provides a streamlined experience for demanding applications, particularly in software engineering where it has demonstrated strong results on technical benchmarks.
A defining characteristic of this model is its recursive self-evolution, where it autonomously performs over 100 iterations to refine its own training process, resulting in measurable performance gains without human intervention. This lineage of self-optimization allows the highspeed variant to achieve output rates of 100 tokens per second, representing a substantial increase in throughput. These advancements make it a practical choice for enterprise-scale deployments that require both rapid response times and high-level reasoning capabilities for complex development workflows.
Quick Info
Powered by- Provider
- ZenMux
- Model key
- minimax/minimax-m2.7-highspeed
- Release date
- Mar 20, 2026
- Last updated
- Mar 20, 2026
- Knowledge cutoff
- 2025-01-01
- AI SDK package
@ai-sdk/anthropic- Input modalities
- Output modalities
- Capabilities
Cost
- Input token cost
- $0.611
- Output token cost
- $2.4439
Limits
- Output tokens
- 131,070 tokens
- Context window
- 204,800 tokens
Latest news about MiniMax M2.7 highspeed
No articles yet. Fetch the latest news to show it here.