Sulat.com
AI models
$10 off the fastest DeepSeek V4.1 Flash, Kimi K3 and GLM 5.3 from Synthetic
Xiaomi logo

Model details

MiMo-V2.5-Pro

MiMo-V2.5-Pro is a sparse Mixture-of-Experts model with over a trillion total parameters and 42 billion active parameters during inference. It inherits a hybrid-attention architecture and three-layer Multi-Token Prediction from the MiMo-V2-Flash backbone, enabling it to handle long-horizon reasoning and complex software engineering challenges. This design makes it particularly strong at agentic workflows and tasks that demand sustained coherence over extended contexts. According to Xiaomi's published benchmarks, MiMo-V2.5-Pro tied for first place among open-weights models on the Artificial Analysis Intelligence Index, demonstrating that efficient expert routing can deliver frontier-level performance without activating the full parameter count on every token.

The model was released alongside the standard MiMo-V2.5 version, both open-sourced under the MIT license for commercial use. XiaomI released both variants with weights available on Hugging Face, enabling enterprises and independent developers to run them locally or in private cloud environments. MiMo-V2.5-Pro excels at agentic tasks such as powering claw-style task systems and benchmark environments like SWE-bench Pro, where it scored competitively against leading proprietary models. With a one-million-token context window and significantly lower operational costs than comparable open-weights alternatives, it targets teams that need high benchmark performance on coding and agentic work without frontier-model pricing.

Xiaomimimo-v2.5-promimo

Quick Info

Powered by
Provider
Xiaomi
Model key
mimo-v2.5-pro
Release date
Apr 22, 2026
Last updated
Jun 24, 2026
Knowledge cutoff
2024-12
Input modalities
Output modalities
Capabilities

Cost

Input token cost
$0.435
Output token cost
$0.87

Limits

Output tokens
131,072 tokens
Context window
1,048,576 tokens

Transparent token rates

Compare MiMo-V2.5-Pro pricing

Rates are shown per one million tokens. Combined means one million input plus one million output tokens.

Browse this family

Latest news about MiMo-V2.5-Pro

Xiaomi

Official sourceDocumentation

Xiaomi's official MiMo updates changelog documents the release of MiMo-V2.5-Pro on 2026-04-23, describing it as a trillion-parameter model with 1T total parameters, 42B activations, and a 1M-token ultra-long context window. The same changelog entry notes that in high-intensity agent scenarios the model performs compara The same MiMo updates page also records that the V2.6 series (V2.6-Pro, V2.6-Flash, and V2.6-Pro-Ultraspeed) launched on 2026-09-22, which supersedes the V2.5-Pro entry above. Because V2.6 is a distinct variant, its details cannot be transferred to V2.5-Pro; only the 2026-04-23 V2.5-Pro changelog line itself constitute

Xiaomi

CoverageRelease Notes

Xiaomi Releases MiMo-V2.5-Pro and MiMo-V2.5: Matching Frontier Model Benchmarks at Significantly Lower Token Cost

ZenMux

Coverage

AIBase reports that on May 29, 2026, Xiaomi issued an end-of-life notice for the MiMo-V2-Pro and MiMo-V2-Omni models, with services scheduled to officially discontinue on June 30, 2026. The notice specifies that mimo-v2-pro will be migrated to mimo-v2.5-pro and mimo-v2-omni will be upgraded to a new model within the mi The retirement of the V2-Pro and V2-Omni variants positions MiMo-V2.5-Pro as the direct successor for Xiaomi's open platform users. Xiaomi framed the transition as encouraging migration to a model with stronger inference capabilities and higher cost-effectiveness, reducing long-term maintenance burden while improving d

Xiaomi

CoverageBenchmark

Analysis of Xiaomi's MiMo-V2.5-Pro and comparison to other AI models across key metrics including quality, price, performance (tokens per second & time to first token), context window & more.

Xiaomi

CoverageBenchmark

See performance metrics across providers for Xiaomi: MiMo-V2.5-Pro - MiMo-V2.5-Pro is Xiaomi’s flagship model, delivering strong performance in general agentic capabilities, complex software engineering, and long-horizon tasks, with top rankings on benchmarks such as ClawEval, GDPVal, and SWE-bench Pro. It can independ

Videos about MiMo-V2.5-Pro

More models around MiMo-V2.5-Pro