Sulat.com
AI models
$10 off the fastest DeepSeek V4.1 Flash, Kimi K3 and GLM 5.3 from Synthetic
EmpirioLabs AI logo

Model details

MiMo V2.5

MiMo V2.5 is positioned as a unified multimodal model that natively handles text, images, audio, and video in a single system, replacing the earlier split between a dedicated agentic model and a separate multimodal variant. According to a late-April 2026 GSMArena report, Xiaomi released the model with open weights, distributing it for download, online use, and through an API, and framed it as offering frontier-level agentic capability suitable for tool-driven and coding-oriented tasks. The same report's description highlights that the model ingests multiple input streams rather than relying on bolted-on adapters, which is a meaningful architectural choice for users building agents that need to interpret screenshots, diagrams, spoken instructions, or video clips alongside text prompts.

For practical deployment, MiMo V2.5 has been made available beyond Xiaomi's own channels through inference platforms such as DeepInfra, where a dedicated integration post was published in July 2026, signaling third-party hosting and broader accessibility for developers who prefer managed inference over running the open weights themselves. The combination of open-weight availability, multimodal understanding across text, images, audio, and video, and an explicit agentic focus makes the model a reasonable fit for teams building assistants that must reason over mixed media, automate multi-step workflows, or prototype tool-calling agents without committing to a closed-weight provider. Teams evaluating it should weigh the open-weight flexibility and multimodal breadth against the need to self-host or choose a compatible inference host.

EmpirioLabs AImimo-v2-5mimo

Quick Info

Powered by
Provider
EmpirioLabs AI
Model key
mimo-v2-5
Release date
Apr 22, 2026
Last updated
Apr 22, 2026
Knowledge cutoff
2024-12
Input modalities
Output modalities
Capabilities

Cost

Input token cost
$0.70
Output token cost
$1.40

Limits

Output tokens
128,000 tokens
Context window
1,000,000 tokens

Transparent token rates

Compare MiMo V2.5 pricing

Rates are shown per one million tokens. Combined means one million input plus one million output tokens.

Browse this family

Latest news about MiMo V2.5

Videos about MiMo V2.5

More models around MiMo V2.5