Sulat.com
AI models
EmpirioLabs AI logo

Model details

MiniMax M2.7 Highspeed

MiniMax M2.7 Highspeed is positioned as a throughput-optimized sibling to the MiniMax-M2.7 flagship large language model, released on March 18, 2026 alongside it. The highspeed variant is described as boosting output speed by approximately 66 percent over the base model, reaching around 100 tokens per second. This focus on faster generation makes the model suitable for latency-sensitive text applications where responsiveness matters more than maximum depth, while still drawing on the same underlying model family.

The model is available through third-party inference platforms such as Runware, where it is listed as a text-generation endpoint supporting synchronous, asynchronous, and streaming delivery modes. Its intended use centers on text-based workloads, and the speed-oriented design suggests practical fit for interactive chat, rapid drafting, and high-volume generation pipelines where quicker token throughput is a priority. As a variant in the broader MiniMax-M2.7 line, it represents a pragmatic choice for users who value output velocity within that model family.

EmpirioLabs AIminimax-m2-7-highspeedminimax

Quick Info

Powered by
Provider
EmpirioLabs AI
Model key
minimax-m2-7-highspeed
Release date
Mar 18, 2026
Last updated
Mar 18, 2026
Input modalities
Output modalities
Capabilities

Cost

Input token cost
$0.30
Output token cost
$1.20

Limits

Output tokens
32,768 tokens
Context window
200,000 tokens

Transparent token rates

Compare MiniMax M2.7 Highspeed pricing

Rates are shown per one million tokens. Combined means one million input plus one million output tokens.

Browse this family

Latest news about MiniMax M2.7 Highspeed

No articles yet. Fetch the latest news to show it here.

Videos about MiniMax M2.7 Highspeed

More models around MiniMax M2.7 Highspeed