Sulat.com
AI models
$10 off the fastest DeepSeek V4.1 Flash, Kimi K3 and GLM 5.3 from Synthetic
LLM Gateway logo

Model details

MiniMax M2.7 Highspeed (MiniMax)

MiniMax M2.7 Highspeed is the latency-optimized variant of the MiniMax M2.7 family, engineered to keep generation snappy even when prompts stretch across long documents or multi-step agent workflows. Its defining design choice is a roughly 200K-token context window paired with a 131K-token maximum output ceiling, giving it room to ingest large codebases, transcripts, or retrieval-augmented payloads without losing continuity between sections. The "highspeed" positioning reflects a deliberate trade toward lower time-to-first-token and higher tokens-per-second throughput rather than a separate architecture, making it a practical default when response time matters more than extended reasoning depth.

Because the weights are openly distributed, the model can be self-hosted for cost-sensitive or latency-critical deployments, but it is also routed through managed endpoints where caching and ephemeral batching lower the effective price of repeated context. It fits naturally into agent and tool-calling pipelines that need to stream long outputs quickly, chat assistants that must stay responsive on bulky system prompts, and developer workflows such as code refactoring or log analysis where large windows of source are read in a single pass. For users who do not need the full context envelope, lighter-weight options in the M2.7 family may be more economical, but for speed at scale, the Highspeed variant is the intended choice.

LLM Gatewayminimax/minimax-m2.7-highspeedminimax

Quick Info

Powered by
Provider
LLM Gateway
Model key
minimax/minimax-m2.7-highspeed
Release date
Mar 18, 2026
Last updated
Mar 18, 2026
Input modalities
Output modalities
Capabilities

Cost

Input token cost
$0.60
Output token cost
$2.40

Limits

Output tokens
131,100 tokens
Context window
204,800 tokens

Transparent token rates

Compare MiniMax M2.7 Highspeed (MiniMax) pricing

Rates are shown per one million tokens. Combined means one million input plus one million output tokens.

Browse this family

Latest news about MiniMax M2.7 Highspeed (MiniMax)

No articles yet. Fetch the latest news to show it here.

Videos about MiniMax M2.7 Highspeed (MiniMax)

More models around MiniMax M2.7 Highspeed (MiniMax)