Model details
MiniMax M2.5 highspeed
MiniMax M2.5 highspeed is engineered as a high-velocity engine for developers and enterprises requiring rapid, intelligent responses. Built upon a Mixture of Experts architecture, the model is designed to maintain the core reasoning depth and robust digital workspace capabilities of the standard M2.5 series. It excels in complex agentic tasks, achieving an 80.2% score on the SWE-Bench Verified benchmark, which positions it as a competitive force in automated coding and software development workflows. By prioritizing architectural efficiency, the model provides a seamless experience for users who need to process large document streams or manage intricate cross-software collaborations without sacrificing logical precision.
The model lineage focuses on rigorous engineering optimization to achieve ultra-low latency, enabling inference speeds that reach up to 100 tokens per second. This performance profile makes it particularly well-suited for latency-sensitive interactive applications and large-scale automated pipelines where real-time responsiveness is critical. Beyond its raw speed, the model demonstrates strong proficiency in tool calling, significantly outperforming legacy benchmarks in functional accuracy. As a forward-looking tool for agentic scenarios, it is designed to integrate deeply into modern productivity environments, providing a scalable solution for high-frequency API calls and complex, multi-step reasoning tasks.
Quick Info
Powered by- Provider
- ZenMux
- Model key
- minimax/minimax-m2.5-lightning
- Release date
- Feb 13, 2026
- Last updated
- Feb 13, 2026
- Knowledge cutoff
- 2025-01-01
- AI SDK package
@ai-sdk/anthropic- Input modalities
- Output modalities
- Capabilities
Cost
- Input token cost
- $0.60
- Output token cost
- $4.80
Limits
- Output tokens
- 131,072 tokens
- Context window
- 204,800 tokens
Latest news about MiniMax M2.5 highspeed
No articles yet. Fetch the latest news to show it here.