Sulat.com
AI models
NaN logo

Model details

MiMo-V2.5

MiMo-V2.5 is positioned as a large omnimodal model in the MiMo family, built as a 310-billion-parameter Mixture-of-Experts architecture with 15 billion active parameters per pass. The NaN model card describes it as natively omnimodal, combining dedicated vision and audio encoders with text input and producing text output, which lets a single deployment handle image and audio prompts alongside natural language. It is served in FP8 quantization, ships under an MIT license that keeps the weights open, and is exposed through an OpenAI-compatible API alongside tool calling and a reasoning mode that the provider recommends running with a generous max-token budget.

The 1M-token context window is the headline capacity feature, supporting long-document reasoning, multi-session agent loops, and retrieval-heavy workloads that would overflow smaller models. Community experiments on consumer and workstation hardware show the practical side of that design: users have run MiMo-V2.5 variants across two and three DGX Spark nodes using tensor parallelism, with a three-node Omni configuration reaching roughly 39 tokens per second at full 1M context using speculative multi-token decoding. That combination of long context, multimodal inputs, open weights, and tool/reasoning hooks makes MiMo-V2.5 a flexible fit for developers building assistant agents, document or media analysis pipelines, and local multi-node research rigs rather than purely chat-style deployments.

NaNmimo-v2.5mimo

Quick Info

Powered by
Provider
NaN
Model key
mimo-v2.5
Release date
Apr 22, 2026
Last updated
Apr 22, 2026
Knowledge cutoff
2024-12
Input modalities
Output modalities
Capabilities

Cost

A provider subscription or plan supersedes token-based pricing for this model.

Limits

Output tokens
131,072 tokens
Context window
1,048,576 tokens

Latest news about MiMo-V2.5

NaN

CoverageRelease Notes

Venice.ai's changelog records "MiMo-V2.5" as a new Text model added on June 11, 2026, with "Privacy TBD" status and availability to all users on the platform. The entry is a single aggregated row that does not provide technical details such as parameter count, architecture, context length, pricing, or benchmark results As evidence, the changelog establishes that a model named MiMo-V2.5 is being served by a third-party inference provider alongside related releases such as GLM 5.2 (Jun 16), Kimi K2.7 Code (Jun 13), and MiniMax M3 Preview (Jun 12), but it does not confirm the upstream developer or attribute the model to the NaN/nan.buil

Videos about MiMo-V2.5

More models around MiMo-V2.5