Currently listed through these providers:
Model details
MiniMax-M2.5-highspeed
MiniMax-M2.5-highspeed belongs to the earlier M2.5 generation of the MiniMax model family, positioned in the documentation as a legacy variant designed for code generation and refactoring tasks. It is documented as offering the same performance profile as the base M2.5 model while delivering significantly faster inference, making it suited to latency-sensitive development workflows such as interactive polyglot coding, precision refactoring, and high-throughput code assistance. As a text-only model aimed at programming-centric use cases, it sits in the lineup alongside related high-speed siblings that share the same performance-with-faster-inference trade-off across the M-series.
Within the broader MiniMax lineup, the high-speed variants are consistently framed as drop-in replacements that preserve capability parity with their base models while prioritizing agility, and the M2.5 generation was originally optimized for complex code generation and refactoring before the introduction of newer M-series successors focused on multimodal coding and recursive self-improvement. This makes M2.5-highspeed a practical fit for teams that want legacy compatibility and faster response times for established code workflows, while still benefiting from the architecture and training emphasis on polyglot programming that characterized the M2.5 line.
Quick Info
Powered by- Provider
- MiniMax (minimaxi.com)
- Model key
- MiniMax-M2.5-highspeed
- Release date
- Feb 13, 2026
- Last updated
- Feb 13, 2026
- Input modalities
- Output modalities
- Capabilities
Cost
- Input token cost
- $0.60
- Output token cost
- $2.40
Limits
- Output tokens
- 131,072 tokens
- Context window
- 204,800 tokens
Latest news about MiniMax-M2.5-highspeed
No articles yet. Fetch the latest news to show it here.