Currently listed through:
Model details
ElevenLabs-v3
ElevenLabs-v3 represents a significant shift in AI-generated audio, moving beyond simple narration to focus on performance-driven speech. Designed to interpret emotional subtext, the model allows for nuanced delivery that captures character, feeling, and timing. By prioritizing the emotional context of a script, it enables more believable character interactions and engaging long-form content, effectively transitioning AI voice technology into the realm of directed voice acting.
The model introduces advanced delivery control through features like Audio Tags, which provide director-level precision over rhythm, pacing, and emphasis. Users can manipulate speech flow by inserting tags for pauses, stammers, or rushed delivery, allowing for dramatic, comedic, or tense performances directly from the script. This capability, combined with a dedicated Text to Dialogue API, offers a versatile toolset for creators looking to shape the cadence and emotional impact of their audio projects with high fidelity.
Quick Info
Powered by- Provider
- Poe
- Model key
- elevenlabs/elevenlabs-v3
- Release date
- Jun 5, 2025
- Last updated
- Jun 5, 2025
- Input modalities
- Output modalities
- Capabilities
Limits
- Output tokens
- 0 tokens
- Context window
- 128,000 tokens