Sulat.com
AI models
Vercel AI Gateway logo

Model details

MiniMax M2.7 High Speed

MiniMax M2.7 High Speed is positioned as a speed-optimized variant of the MiniMax M2.7 family, explicitly marketed as delivering the same underlying performance with "faster and more agile" output at roughly 100 tokens per second. The model is exposed through the Vercel AI Gateway and listed on the AI SDK Playground under the MiniMax provider umbrella, and it also appears via a third-party route on EmpirioLabs AI, suggesting it can be reached through more than one upstream endpoint while keeping the same underlying weights. That dual availability makes it attractive for teams that want to standardize on a single model identity but negotiate latency or capacity across providers.

In practical terms, the variant suits latency-sensitive workloads where MiniMax M2.7's quality is already acceptable: interactive assistants, multi-step agent loops, retrieval-augmented chat, and developer tooling that streams long completions. The 204,800-token context window leaves room for sizable system prompts, tool definitions, and retrieved documents, so high-speed generation does not come at the cost of a shrunken working memory. Because the framing emphasizes matching baseline M2.7 quality while shifting the speed/responsiveness dial upward rather than retraining for a different capability profile, it slots in cleanly as a drop-in "fast" preset inside the broader M2.7 lineup.

Vercel AI Gatewayminimax/minimax-m2.7-highspeedminimax

Quick Info

Powered by
Provider
Vercel AI Gateway
Model key
minimax/minimax-m2.7-highspeed
Release date
Mar 18, 2026
Last updated
Mar 18, 2026
Input modalities
Output modalities
Capabilities

Cost

Input token cost
$0.60
Output token cost
$2.40

Limits

Output tokens
131,100 tokens
Context window
204,800 tokens

Latest news about MiniMax M2.7 High Speed

Videos about MiniMax M2.7 High Speed

More models around MiniMax M2.7 High Speed