Currently listed through these providers:
Model details
MiniMax M2.7 Highspeed
MiniMax M2.7 Highspeed is positioned as a throughput-optimized sibling to the MiniMax-M2.7 flagship large language model, released on March 18, 2026 alongside it. The highspeed variant is described as boosting output speed by approximately 66 percent over the base model, reaching around 100 tokens per second. This focus on faster generation makes the model suitable for latency-sensitive text applications where responsiveness matters more than maximum depth, while still drawing on the same underlying model family.
The model is available through third-party inference platforms such as Runware, where it is listed as a text-generation endpoint supporting synchronous, asynchronous, and streaming delivery modes. Its intended use centers on text-based workloads, and the speed-oriented design suggests practical fit for interactive chat, rapid drafting, and high-volume generation pipelines where quicker token throughput is a priority. As a variant in the broader MiniMax-M2.7 line, it represents a pragmatic choice for users who value output velocity within that model family.
Quick Info
Powered by- Provider
- EmpirioLabs AI
- Model key
- minimax-m2-7-highspeed
- Release date
- Mar 18, 2026
- Last updated
- Mar 18, 2026
- Input modalities
- Output modalities
- Capabilities
Cost
- Input token cost
- $0.30
- Output token cost
- $1.20
Limits
- Output tokens
- 32,768 tokens
- Context window
- 200,000 tokens
Transparent token rates
Compare MiniMax M2.7 Highspeed pricing
Rates are shown per one million tokens. Combined means one million input plus one million output tokens.
Latest news about MiniMax M2.7 Highspeed
No articles yet. Fetch the latest news to show it here.
Videos about MiniMax M2.7 Highspeed
More models around MiniMax M2.7 Highspeed
This exact model name is also listed by 11 other providers.