Sulat.com
AI models
$10 off the fastest DeepSeek V4.1 Flash, Kimi K3 and GLM 5.3 from Synthetic
NovitaAI logo

Model details

XiaomiMiMo/MiMo-V2-Flash

MiMo-V2-Flash is a Mixture-of-Experts language model built to balance massive scale with efficient execution. It features 309 billion total parameters while utilizing only 15 billion active parameters per forward pass, a design choice aimed at delivering high-speed reasoning. The architecture incorporates a hybrid attention mechanism that interleaves sliding window attention with global attention, which helps reduce the KV cache footprint significantly. By integrating Multi-Token Prediction, the model achieves faster inference speeds, making it well-suited for agentic workflows that require both depth and responsiveness.

The model was pre-trained on 27 trillion tokens and utilizes a Multi-Teacher On-Policy Distillation paradigm to refine its capabilities. This post-training method allows the model to learn from domain-specialized teachers, effectively absorbing expert knowledge to improve performance on complex tasks. With its ability to handle extended context lengths, the model is positioned to compete with frontier-level systems in coding and reasoning benchmarks. Its open-source nature and technical design make it a versatile tool for developers looking to integrate high-capacity, efficient AI into their own infrastructure.

NovitaAIxiaomimimo/mimo-v2-flashmimo

Quick Info

Powered by
Provider
NovitaAI
Model key
xiaomimimo/mimo-v2-flash
Release date
Dec 19, 2025
Last updated
Dec 19, 2025
Knowledge cutoff
2024-12
Input modalities
Output modalities
Capabilities

Cost

Input token cost
$0.10
Output token cost
$0.30

Limits

Output tokens
32,000 tokens
Context window
262,144 tokens

Transparent token rates

Compare XiaomiMiMo/MiMo-V2-Flash pricing

Rates are shown per one million tokens. Combined means one million input plus one million output tokens.

Browse this family

Latest news about XiaomiMiMo/MiMo-V2-Flash

NovitaAI

Coverage

Xiaomi's open-source large model MiMo-V2-Flash API has launched a payment function and is about to start a paid model. Users can log in to their accounts to obt

NovitaAI

CoverageDiscourse

Hi! I'm Niels, part of the community science team at Hugging Face.

NovitaAI

Coverage

Xiaomi announced that the free public testing period of its self-developed large model MiMo-V2-Flash has been extended by 20 days, until January 20, 2026. The m

NovitaAI

Coverage

Xiaomi unveils MiMo-V2-Flash open-source AI model, signaling strategic AI push.

Videos about XiaomiMiMo/MiMo-V2-Flash

More models around XiaomiMiMo/MiMo-V2-Flash