Sulat.com
AI models
$10 off the fastest DeepSeek V4.1 Flash, Kimi K3 and GLM 5.3 from Synthetic
ZenMux logo

Model details

MiMo-V2-Flash

MiMo-V2-Flash is built on a sparse Mixture-of-Experts architecture that utilizes 309 billion total parameters, with 15 billion active parameters per inference. To balance computational efficiency with high-level performance, the model employs a hybrid attention mechanism that interleaves a 128-token sliding window with full attention in a 5:1 ratio. This design is specifically engineered to support advanced reasoning, software engineering, and tool invocation, positioning the model as a foundational technology for complex agentic tasks and interconnected device ecosystems.

The development of this model prioritizes robust agent execution and collaborative reasoning, supported by stable reinforcement learning techniques during post-training. By focusing on these areas, the model achieves top-tier results on benchmarks such as the AIME 2025 math competition and the GPQA-Diamond scientific knowledge test, while also leading in software engineering evaluations like SWE-bench. These capabilities reflect a strategic shift toward models that can interact effectively with physical and digital environments, making it a strong candidate for developers building sophisticated, multi-step autonomous systems.

ZenMuxxiaomi/mimo-v2-flashmimo

Quick Info

Powered by
Provider
ZenMux
Model key
xiaomi/mimo-v2-flash
Release date
Dec 16, 2025
Last updated
Feb 4, 2026
Knowledge cutoff
2024-12-01
Input modalities
Output modalities
Capabilities

Cost

Input token cost
$0.10
Output token cost
$0.30

Limits

Output tokens
65,536 tokens
Context window
262,144 tokens

Transparent token rates

Compare MiMo-V2-Flash pricing

Rates are shown per one million tokens. Combined means one million input plus one million output tokens.

Browse this family

Latest news about MiMo-V2-Flash

ZenMux

CoverageRelease Notes

Opper AI's Xiaomi release tracker provides dated timeline and pricing context for the MiMo family, listing MiMo V2 Flash with a 16 December 2025 release date, a 262K context window, and pricing of $0.10 per million input tokens and $0.30 per million output tokens, alongside an Artificial Analysis intelligence score of This release-tracker view clarifies the chronological progression of Xiaomi's open-source MoE lineup, showing MiMo V2 Flash as the December 2025 entry point before the larger 1M-context V2 Pro and V2.5 Pro models followed in March and April 2026. A minor discrepancy exists: Opper lists the Flash context at 262K while Z

Videos about MiMo-V2-Flash

More models around MiMo-V2-Flash