Currently listed through these providers:
Model details
MiMo-V2.5-TTS
MiMo-V2.5-TTS sits inside Xiaomi's MiMo V2.5 family as a dedicated text-to-speech engine, accepting written input and producing audio output. Independent catalog listings describe it as an active speech-synthesis model whose standout trait is voice design, meaning callers can shape the character and delivery of the generated speech rather than relying on a single fixed voice. This focus on controllable synthesis positions the model for product teams that need branded, expressive, or multi-voice narration without training their own acoustic models, and the variant naming seen across aggregators suggests a flexible endpoint aimed at customizable speech output.
In practice, MiMo-V2.5-TTS is best suited for developers and content workflows that want on-demand voice generation with stylistic control, such as audiobooks, virtual assistants, accessibility narration, or localized voiceover pipelines. Third-party reseller listings show the model is being distributed through a range of API providers at widely varying rates, which is worth keeping in mind when comparing access options. As part of the broader MiMo V2.5 lineup, the model extends Xiaomi's portfolio beyond pure language understanding into generative audio, signaling continued investment in speech synthesis alongside the family's text-oriented offerings.
Quick Info
Powered by- Provider
- Xiaomi Token Plan (Singapore)
- Model key
- mimo-v2.5-tts
- Release date
- Apr 22, 2026
- Last updated
- Apr 22, 2026
- Input modalities
- Output modalities
- Capabilities
Cost
A provider subscription or plan supersedes token-based pricing for this model.
Limits
- Output tokens
- 8,192 tokens
- Context window
- 8,192 tokens
Latest news about MiMo-V2.5-TTS
Videos about MiMo-V2.5-TTS
More models around MiMo-V2.5-TTS
This exact model name is also listed by 2 other providers.