Sulat.com
AI models
$10 off the fastest DeepSeek V4.1 Flash, Kimi K3 and GLM 5.3 from Synthetic
Deep Infra logo

Model details

MiMo-V2.6-Pro

MiMo-V2.6-Pro is the flagship release in Xiaomi's MiMo-V2.6 series, described by the company as its most capable model to date and built as a natively omnimodal system that handles text along with image, audio, and video inputs in a single architecture. The series was released and open-sourced on September 22nd, 2026, marking a key step in Xiaomi's exploration of a recursive self-improvement path in which reinforcement learning compute is scaled on verifiable, complex tasks so the model can extend its capability frontier through exploration and feedback. Alongside the Pro model, Xiaomi introduced a Flash variant aimed at balancing intelligence, efficiency, and cost, as well as a Pro-UltraSpeed variant that targets up to 20x faster output at comparable quality for latency-sensitive deployments.

Independent benchmark reporting from Xiaomi positions MiMo-V2.6-Pro as a strong contender in agentic and coding evaluations, with scores of 71.9 on DeepSWE v1.1, 26.5 on ProgramBench, and 63.2 on the in-house MiMo Code Bench, results that place it ahead of the Flash sibling and well above the prior MiMo-V2.5-Pro generation. The model is designed for practical agentic workflows rather than narrow chat use, with reported performance on additional agent benchmarks such as Toolathlon-verified, GDPVal 2.1 AA, Automation Bench v1.0.6, and Agents' Last Exam reinforcing a focus on tool use, automation, and long-horizon task completion. Open-weight availability makes MiMo-V2.6-Pro well suited to teams that want to self-host a frontier-grade omnimodal model for coding assistants, research agents, or complex multi-step automation pipelines.

Deep InfraXiaomiMiMo/MiMo-V2.6-Promimo

Quick Info

Powered by
Provider
Deep Infra
Model key
XiaomiMiMo/MiMo-V2.6-Pro
Release date
Sep 22, 2026
Last updated
Sep 22, 2026
Input modalities
Output modalities
Capabilities

Cost

Input token cost
$0.43
Output token cost
$0.87

Limits

Output tokens
131,072 tokens
Context window
1,048,576 tokens

Transparent token rates

Compare MiMo-V2.6-Pro pricing

Rates are shown per one million tokens. Combined means one million input plus one million output tokens.

Browse this family

Latest news about MiMo-V2.6-Pro

Deep Infra

CoverageBenchmark

Xiaomi released MiMo V2.6 on September 22, 2026, a family of open-weight, natively omnimodal language models that handle text, images, video, and audio with a 1 million-token context window. Built around the slogan "Scaling Reinforcement Learning Toward Self-Improvement," the lineup includes Pro, Flash, Pro UltraSpeed, and a 9B distill, all under an MIT license and available on Hugging Face, ModelScope, Xiaomi's API, and OpenRouter. Pro is a 1.02-trillion-parameter sparse MoE model with 42B active parameters, Flash is a 309B/15B-active variant, and UltraSpeed targets low latency. Xiaomi reports Pro leading peers on AutomationBench v1.0.6 (53.1), Terminal Bench 2.1 (89.9), and CyberGym (94.0), while trailing on harder coding and exploit benchmarks like Terminal Bench 4.0 (34.9). Artificial Analysis places Pro at an Intelligence Index of 46 at roughly $0.13 per task, with OpenRouter pricing of $0.14/$0.28 per million tokens for Flash and $0.435/$0.87 for Pro.

Deep Infra

CoverageBenchmark

On 22 September 2026, Xiaomi released its MiMo-V2.6 model family as open-weight checkpoints under an MIT licence, with the flagship MiMo-V2.6-Pro topping Artificial Analysis's open-weight Intelligence Index at a score of 46. The release comes from Xiaomi's MiMo team and bundles four models together with a technical report and open-source RL tooling on GitHub. Xiaomi frames the launch around a single mixed reinforcement-learning run across coding, agents, visual tasks and cybersecurity, replacing earlier domain-specific training. MiMo-V2.6-Pro is a 1.02T-parameter sparse MoE with 42B active parameters, a 1M-token context, and text, image, video and audio input, with a 573 GB checkpoint. API pricing sits at $0.435 per million input tokens and $0.87 per million output tokens, while the Flash sibling (309B total, 15B active) is priced at $0.14 and $0.28, and an UltraSpeed variant costs ten times more. Self-hosting Pro requires datacentre GPUs such as 8x H200, whereas Flash fits on a single multi-GPU node and a 9B Qwen distill runs on a laptop. Xiaomi was named in Anthropic's 10 September 2026 distillation report alongside six other China-based labs, though no link to V2.6 has been documented.

Videos about MiMo-V2.6-Pro

More models around MiMo-V2.6-Pro