Sulat.com
AI models
$10 off the fastest DeepSeek V4.1 Flash, Kimi K3 and GLM 5.3 from Synthetic
CrossModel logo

Model details

MiMo-V2.6-Pro

MiMo-V2.6-Pro is Xiaomi's flagship reasoning model in the newly released V2.6 series, built as a natively omnimodal system that natively takes in text, images, audio, and video for unified understanding. Xiaomi frames the series as a step along an RSI (Recursively Self-Improving) training path, scaling reinforcement-learning compute on verifiable, complex tasks so the model can expand its capability frontier through exploration and feedback rather than relying on static pretraining alone. The architecture is positioned at trillion-parameter scale, reflecting Xiaomi's push into very large reasoning models aimed at long-context, multi-step problem solving.

In Xiaomi's own benchmark reporting, MiMo-V2.6-Pro is shown as a top-tier coding and agentic model, posting a 71.9 score on DeepSWE v1.1 software-engineering tasks and a 63.2 score on Xiaomi's in-house MiMo Code Bench, while also reaching 26.5 on ProgramBench and competitive results on agent and automation evaluations such as GDPVal, Toolathlon-verified, Automation Bench, and Agents' Last Exam. Practically, it is aimed at complex projects, long-horizon work, cybersecurity, and research workloads where deep reasoning and tool use matter more than lightweight chat, and it sits alongside a faster Flash variant and a Pro-UltraSpeed deployment tuned for very low-latency generation.

CrossModelxiaomi/mimo-v2.6-promimo

Quick Info

Powered by
Provider
CrossModel
Model key
xiaomi/mimo-v2.6-pro
Release date
Sep 22, 2026
Last updated
Sep 22, 2026
Input modalities
Output modalities
Capabilities

Cost

Input token cost
$0.47
Output token cost
$0.94

Limits

Output tokens
131,072 tokens
Context window
1,048,576 tokens

Transparent token rates

Compare MiMo-V2.6-Pro pricing

Rates are shown per one million tokens. Combined means one million input plus one million output tokens.

Browse this family

Latest news about MiMo-V2.6-Pro

Xiaomi

CoverageBenchmark

Forkast reports that MiMo-V2.6-Pro uses a frozen-router Mixture-of-Experts architecture with 1.02 trillion total parameters and 42 billion active parameters, incorporating hybrid attention mechanisms and a 5-layer MTP speculative decoder. Training used fully asynchronous GRPO with 1,568 prompts across 16 rollouts per s On vendor-reported benchmarks, MiMo-V2.6-Pro achieved 71.9% on DeepSWE v1.1, 89.9% on Terminal Bench 2.1 (surpassing Claude Opus 5 and GPT-5.6 Sol), 94.0% on CyberGym (leading all reported models), and 53.1% on AutomationBench v1.0.6. The article also contextualizes the release alongside the Alibaba V900 chip unveiling

Xiaomi

Official sourceDocumentation

Xiaomi officially released and open-sourced the MiMo-V2.6 series, comprising two native fully multimodal models (Pro and Flash), on September 22, 2026. The release documents a 6-day live RL training phase as part of Xiaomi's recursive self-improvement exploration, building on verifiable complex tasks and scaling up RL The V2.6 series adopts the same API pricing as V2.5, with Xiaomi framing MiMo-V2.6-Pro as setting a new cost-performance record among domestic large language models at 1/20 to 1/60 the price of overseas models at comparable intelligence. MiMo-V2.6-Pro is described as one of the domestic open-source models with the larg

Videos about MiMo-V2.6-Pro

More models around MiMo-V2.6-Pro