Sulat.com
AI models
$10 off the fastest DeepSeek V4.1 Flash, Kimi K3 and GLM 5.3 from Synthetic
Google logo

Model details

Gemini 2.5 Pro Preview TTS

Gemini 2.5 Pro Preview TTS is designed to bridge the gap between high-fidelity audio output and precise developer control. As part of the Gemini family, this model is built to handle complex text-to-speech tasks by interpreting nuanced instructions directly within the prompt. Its architecture is optimized for applications that require more than just standard narration, allowing users to generate speech that conveys specific emotional states and tones, making it a versatile tool for creating expressive and lifelike synthetic voices.

The model distinguishes itself through its ability to process emotional cues and SSML tags, enabling developers to inject specific vocalizations like excitement, sarcasm, or empathy into their audio output. By supporting these granular controls, the model provides a high level of customization that is often missing in more basic text-to-speech solutions. This capability makes it a strong candidate for projects requiring sophisticated audio synthesis, offering a balance of performance and creative flexibility that supports advanced, tone-aware media generation.

Googlegemini-2.5-pro-preview-ttsgemini-flash

Quick Info

Powered by
Provider
Google
Model key
gemini-2.5-pro-preview-tts
Release date
May 1, 2025
Last updated
May 1, 2025
Knowledge cutoff
2025-01
Input modalities
Output modalities
Capabilities

Cost

Input token cost
$1.00
Output token cost
$20.00

Limits

Output tokens
16,384 tokens
Context window
8,192 tokens

Latest news about Gemini 2.5 Pro Preview TTS

Google

CoveragePreview

Calculate the cost of using gemini-2.5-pro-preview-tts from Google Gemini for Chat workloads. Input: $1.25 per 1M tokens, Output: $10.00 per 1M tokens

Videos about Gemini 2.5 Pro Preview TTS

More models around Gemini 2.5 Pro Preview TTS