Sulat.com
AI models
$10 off the fastest DeepSeek V4.1 Flash, Kimi K3 and GLM 5.3 from Synthetic
Vercel AI Gateway logo

Model details

Gemini 3.8 Flash TTS

Gemini 3.8 Flash TTS is Google's text-to-speech model designed for deep creative direction and character voice design across gaming, immersive audiobooks, podcasts, and interactive media. Launched by Google alongside a lighter sibling variant, it is described as one of the company's most expressive audio generation models to date, reflecting continued investment in natural and controllable speech synthesis for narrative and entertainment applications.

According to the Gemini 3.8 Audio model card, the TTS variants are based on Gemini 3 Pro and accept text input up to 8K tokens while returning audio output up to 64K tokens, giving creators substantial room for long-form script-to-voice workflows. The model supports voice design and replication capabilities and, together with its Lite counterpart, ranked first and second worldwide on Hume AI's Overall Quality Index, underscoring its standing in third-party expressive-voice benchmarks for creators who need fine-grained control over tone, pacing, and character.

Vercel AI Gatewaygoogle/gemini-3.8-flash-ttsgemini

Quick Info

Powered by
Provider
Vercel AI Gateway
Model key
google/gemini-3.8-flash-tts
Release date
Sep 23, 2026
Last updated
Sep 23, 2026
Input modalities
Output modalities
Capabilities
Base catalog fields only

Cost

Input token cost
$0.50
Output token cost
$9.00

Limits

Output tokens
0 tokens
Context window
0 tokens

Latest news about Gemini 3.8 Flash TTS

Vercel AI Gateway

Official sourceOfficial

On September 23, 2026, Google announced Gemini 3.8 Flash TTS (alongside the sibling Gemini 3.8 Flash-Lite TTS) as new additions to the Gemini text-to-speech model family, positioning them as the company's most expressive audio generation models to date. Flash TTS is specifically built for deep creative direction and ch The announcement also highlights workflow and safety features: built-in SynthID watermarking for generated audio, support for multi-speaker scenes, line-by-line stage directions, and integration into Google's broader Gemini product surface for both developers and enterprise users. Google describes the release as a shif

Videos about Gemini 3.8 Flash TTS

More models around Gemini 3.8 Flash TTS