OpenRouter
MiniMax, a Chinese AI startup, has released its MiniMax-01 family of open-source models. The company says its MiniMax-Text-01 can handle contexts up to 4 million tokens - double the capacity of its closest competitor.
Model details
The MiniMax-01 series is built upon an innovative architecture centered on Lightning Attention, a mechanism designed to achieve near-linear computational complexity. This design allows the models to handle vast amounts of information, effectively enabling AI agents to maintain long-term memory by processing extensive inputs in a single exchange. To manage this scale, the architecture integrates a Mixture of Experts framework featuring 32 experts and a total of 456 billion parameters, with 45.9 billion parameters activated for each token. This structural approach ensures that the models maintain high performance while managing the computational demands of ultra-long sequences.
The development of this series involved highly efficient parallel strategies and computation-communication overlap techniques to support training and inference at a massive scale. The vision-language variant was further refined through continued training on 512 billion tokens, allowing it to bridge text and visual data processing. By combining these methods with their specialized attention mechanism, the models demonstrate performance comparable to top-tier industry standards. This technical foundation positions the series as a robust solution for tasks requiring deep analysis of extensive datasets, offering a scalable path forward for developers building complex, context-heavy agentic systems.
Transparent token rates
Rates are shown per one million tokens. Combined means one million input plus one million output tokens.
OpenRouter
MiniMax, a Chinese AI startup, has released its MiniMax-01 family of open-source models. The company says its MiniMax-Text-01 can handle contexts up to 4 million tokens - double the capacity of its closest competitor.
OpenRouter
It is 10x cheaper than GPT-4o, with API costs at $0.20/M input tokens and $1.10/M output tokens.
OpenRouter
MiniMax launches open-source MiniMax-01 series with Text-01 & VL-01 models. Lightning Attention architecture, 456B parameters, 4M token context, 20-32x longer than competitors. Starting $0.2/M tokens.