Currently listed through these providers:
Model details
MiMo V2.5 Pro
MiMo V2.5 Pro is Xiaomi's flagship text-only reasoning model, built around a trillion-parameter sparse design with roughly 42 billion active parameters per forward pass. Xiaomi's release notes describe it as an "efficient architecture" tuned for an ultra-long one-million-token context window, positioning it as a backbone for agent loops that need to keep very large prompts, tool traces, and retrieved material in working memory without losing coherence across long sessions.
In Xiaomi's own framing, the model is aimed squarely at high-intensity agent scenarios, with the launch notes claiming performance comparable to Claude Opus 4.6 on that workload. Above.dev exposes it as an "UltraSpeed" variant that peaks around 1,000 tokens per second through an OpenAI-compatible endpoint, and the same route is advertised as passing upstream cache discounts straight through. That combination of long context, agent-oriented reasoning, and high-throughput serving makes it a practical fit for autonomous pipelines that re-bill the same prompt across many tool calls, where cached-token economics and steady token generation directly shape real-world cost and latency.
Quick Info
Powered by- Provider
- above.dev
- Model key
- mimo-v2.5-pro
- Release date
- Apr 22, 2026
- Last updated
- Apr 22, 2026
- Knowledge cutoff
- 2024-12
- Input modalities
- Output modalities
- Capabilities
Cost
- Input token cost
- $0.5077
- Output token cost
- $1.0154
Limits
- Output tokens
- 131,072 tokens
- Context window
- 1,048,576 tokens
Latest news about MiMo V2.5 Pro
Videos about MiMo V2.5 Pro
More models around MiMo V2.5 Pro
This exact model name is also listed by 25 other providers.