Sulat.com
AI models
$10 off the fastest DeepSeek V4.1 Flash, Kimi K3 and GLM 5.3 from Synthetic
Xiaomi Token Plan (Europe) logo

Model details

MiMo-V2.5-TTS-VoiceClone

MiMo-V2.5-TTS-VoiceClone sits within Xiaomi's MiMo-V2.5-TTS Series, a speech synthesis family that converts input text into natural, fluent speech and gives developers fine-grained control over how that speech sounds. According to the official MiMo documentation, the broader series supports out-of-the-box built-in voices, voice design driven by text descriptions, voice replication from audio samples, and diverse stylistic controls covering speed, emotion, role-play, and dialects. Within this lineup, the VoiceClone variant is positioned specifically for replication, allowing arbitrary voices to be reconstructed from sample audio rather than relying on presets.

The V2.5 series represents the current generation of Xiaomi's speech synthesis offerings, replacing the earlier MiMo-V2 line, which was deprecated on June 30 with users encouraged to migrate. Low-latency streaming output has been restored for the family, returning audio in real time, which makes the models well suited to interactive applications such as conversational agents, accessibility tools, and localized voice interfaces. Practical strengths therefore center on expressive, style-aware synthesis and the ability to recreate a target voice from a sample, while fitting naturally into projects that already use other members of the MiMo ecosystem.

Xiaomi Token Plan (Europe)mimo-v2.5-tts-voiceclonemimo

Quick Info

Powered by
Provider
Xiaomi Token Plan (Europe)
Model key
mimo-v2.5-tts-voiceclone
Release date
Apr 22, 2026
Last updated
Apr 22, 2026
Input modalities
Output modalities
Capabilities

Cost

A provider subscription or plan supersedes token-based pricing for this model.

Limits

Output tokens
8,192 tokens
Context window
8,192 tokens

Latest news about MiMo-V2.5-TTS-VoiceClone

Xiaomi Token Plan (Europe)

Official sourceDocumentation

Xiaomi's official MiMo platform changelog (mimo.mi.com) records that on 2026-4-23, five new models were added to all Token Plan tiers, explicitly listing mimo-v2.5-tts-voiceclone alongside mimo-v2.5-tts, mimo-v2.5-tts-voicedesign, mimo-v2.5, and mimo-v2.5-pro. The same update unified token consumption across context le The changelog further enumerates time-limited promotions layered onto the plan — a returning-user usage-reset bonus, a 23% first-time auto-renewal discount for new subscribers (30% for returning users enabling auto-renewal), a 12% saving on annual plans versus monthly auto-renewal, and an off-peak rate of 0.8x token co

Xiaomi Token Plan (Europe)

Coverage

Aibase reports Xiaomi's launch of the MiMo-V2.5 Full-Stack Speech Model Series, comprising three TTS models and one open-source ASR model covering voice input and output for the agent era. The report frames TTS as shifting toward a "language as control" paradigm where natural-language descriptions direct performance pa The article explicitly names MiMo-V2.5-TTS-VoiceClone and describes it as replicating a target voice with a small sample (e.g., 30 seconds of audio) while retaining style instruction and audio tag response, with use cases including virtual anchors and personalized assistants. It also covers the layered script mechanism

Videos about MiMo-V2.5-TTS-VoiceClone

More models around MiMo-V2.5-TTS-VoiceClone