Currently listed through:
Model details
Xiaomi MiMo-V2.5
MiMo-V2.5 is built for long, tool-driven work that may mix software code with visual, video, and audio information. Its sparse design activates 15 billion of 310 billion parameters per token, seeking a balance between broad model capacity and practical inference efficiency. A three-layer multi-token prediction module can accelerate decoding, while EAGLE speculative decoding can use the accompanying MTP weights.
The architecture interleaves sliding-window and global attention to limit key-value cache growth while retaining access across long inputs. This foundation suits complex coding, multistep reasoning, document analysis, and agent workflows that must interpret multiple media types. Its native multimodal path uses a vision encoder and audio transformer, and the documented deployment provides multimodal understanding through an OpenAI-compatible interface.
Quick Info
Powered by- Provider
- AIHubMix
- Model key
- xiaomi-mimo-v2.5
- Release date
- Apr 22, 2026
- Last updated
- May 13, 2026
- Knowledge cutoff
- 2024-12
- Input modalities
- Output modalities
- Capabilities
Cost
- Input token cost
- $0.44
- Output token cost
- $2.20
Limits
- Output tokens
- 131,072 tokens
- Context window
- 1,048,576 tokens