Sulat.com
AI models
$10 off the fastest DeepSeek V4.1 Flash, Kimi K3 and GLM 5.3 from Synthetic
above.dev logo

Model details

MiMo V2.6 Flash

MiMo V2.6 Flash belongs to Xiaomi's MiMo family of language models, a line that Xiaomi has been advancing through reinforcement-learning experiments after the MiMo V2.5 open-source release in April 2026. The V2.6 generation is positioned around the question of how far reinforcement learning can scale, with MiMo V2.6 Flash designed as the lighter, faster counterpart to the Pro variant. Third-party tracking places the model at rank 31 on the LLM Stats composite scoreboard, with its strongest showing in coding, where it sits in the top 10 percent at position 22 of 273 tracked models. That coding tilt, paired with the Flash suffix, suggests a design intent aimed at developers who need quick responses on programming and tool-use tasks rather than long-form chat.

On capability breadth, MiMo V2.6 Flash earns an average tier for tool calling and reasoning on the same third-party leaderboard, while landing below the top half for chat and vision, which points to a model that is more comfortable with structured, task-oriented prompts than with open-ended conversation or image understanding. LLM Stats' cost-efficiency chart lists MiMo V2.6 Flash at roughly $0.15 per million blended tokens, sitting between DeepSeek V4 Flash variants and indicating a focus on affordable inference. Because Xiaomi has not yet shipped an official model card or release notes for this version, the practical picture comes mainly from aggregator data; once Xiaomi publishes details, the model's exact training recipe, supported context length, and intended deployment scenarios should become clearer for teams evaluating it for coding assistants and lightweight agent workflows.

above.devmimo-v2.6-flashmimo

Quick Info

Powered by
Provider
above.dev
Model key
mimo-v2.6-flash
Release date
Sep 22, 2026
Last updated
Sep 22, 2026
Input modalities
Output modalities
Capabilities

Cost

Input token cost
$0.1692
Output token cost
$0.3385

Limits

Output tokens
131,072 tokens
Context window
1,048,576 tokens

Transparent token rates

Compare MiMo V2.6 Flash pricing

Rates are shown per one million tokens. Combined means one million input plus one million output tokens.

Browse this family

Latest news about MiMo V2.6 Flash

above.dev

Coverage

Times of AI reported on September 22, 2026 that Xiaomi announced the release and open-sourcing of the MiMo-V2.6 series, which includes MiMo-V2.6-Pro and MiMo-V2.6-Flash alongside a Pro-UltraSpeed variant created for swift generation. The article notes Xiaomi is releasing model weights, a technical report, training envi The coverage centers on the MiMo-V2.6-Pro checkpoint, reporting a 1.02 trillion total parameter count with 42 billion activated per token, a 70-layer backbone with 60 sliding-window attention layers and 10 global-attention layers, and 384 routed experts with eight active per token. While Flash-specific quantitative cla

above.dev

Coverage

VentureBeat reported on September 21, 2026 that Xiaomi released MiMo-V2.6-Pro and MiMo-V2.6-Flash as open-weight, MIT-licensed models available on Hugging Face. The article explicitly describes MiMo-V2.6-Flash as a smaller, substantially cheaper model that retains the same 1-million-token context window and native mult The article provides Flash-specific pricing: $0.14 per million uncached input tokens and $0.28 per million output tokens, compared with $0.435 and $0.87 respectively for Pro, with Pro measured by Artificial Analysis at roughly 134 tokens per second output speed. It also notes Xiaomi's API supports text, image, audio an

above.dev

Coverage

On September 22, 2026, Xiaomi released and open-sourced the MiMo-V2.6 series, which explicitly includes the MiMo-V2.6-Flash variant. According to the official launch page, MiMo-V2.6-Flash is one of two natively omnimodal models in the series and is positioned as striking the best balance between intelligence, efficienc The launch page frames the MiMo-V2.6 series as part of Xiaomi's exploration of the RSI path — scaling RL compute on verifiable, complex tasks — and notes that a MiMo-V2.6-Pro-UltraSpeed variant is being rolled out for up to 20x faster output at the same quality. The series as a whole is presented as natively omnimodal

Videos about MiMo V2.6 Flash

More models around MiMo V2.6 Flash