MiMo-V2.6-Pro is Xiaomi's flagship reasoning model in the newly released V2.6 series, built as a natively omnimodal system that natively takes in text, images, audio, and video for unified understanding. Xiaomi frames the series as a step along an RSI (Recursively Self-Improving) training path, scaling reinforcement-learning compute on verifiable, complex tasks so the model can expand its capability frontier through exploration and feedback rather than relying on static pretraining alone. The architecture is positioned at trillion-parameter scale, reflecting Xiaomi's push into very large reasoning models aimed at long-context, multi-step problem solving.
In Xiaomi's own benchmark reporting, MiMo-V2.6-Pro is shown as a top-tier coding and agentic model, posting a 71.9 score on DeepSWE v1.1 software-engineering tasks and a 63.2 score on Xiaomi's in-house MiMo Code Bench, while also reaching 26.5 on ProgramBench and competitive results on agent and automation evaluations such as GDPVal, Toolathlon-verified, Automation Bench, and Agents' Last Exam. Practically, it is aimed at complex projects, long-horizon work, cybersecurity, and research workloads where deep reasoning and tool use matter more than lightweight chat, and it sits alongside a faster Flash variant and a Pro-UltraSpeed deployment tuned for very low-latency generation.