Currently listed through these providers:
Model details
MiniMax M2.7 High Speed
MiniMax M2.7 High Speed is positioned as a speed-optimized variant of the MiniMax M2.7 family, explicitly marketed as delivering the same underlying performance with "faster and more agile" output at roughly 100 tokens per second. The model is exposed through the Vercel AI Gateway and listed on the AI SDK Playground under the MiniMax provider umbrella, and it also appears via a third-party route on EmpirioLabs AI, suggesting it can be reached through more than one upstream endpoint while keeping the same underlying weights. That dual availability makes it attractive for teams that want to standardize on a single model identity but negotiate latency or capacity across providers.
In practical terms, the variant suits latency-sensitive workloads where MiniMax M2.7's quality is already acceptable: interactive assistants, multi-step agent loops, retrieval-augmented chat, and developer tooling that streams long completions. The 204,800-token context window leaves room for sizable system prompts, tool definitions, and retrieved documents, so high-speed generation does not come at the cost of a shrunken working memory. Because the framing emphasizes matching baseline M2.7 quality while shifting the speed/responsiveness dial upward rather than retraining for a different capability profile, it slots in cleanly as a drop-in "fast" preset inside the broader M2.7 lineup.
Quick Info
Powered by- Provider
- Vercel AI Gateway
- Model key
- minimax/minimax-m2.7-highspeed
- Release date
- Mar 18, 2026
- Last updated
- Mar 18, 2026
- Input modalities
- Output modalities
- Capabilities
Cost
- Input token cost
- $0.60
- Output token cost
- $2.40
Limits
- Output tokens
- 131,100 tokens
- Context window
- 204,800 tokens