Vercel AI Gateway
Google has launched two new text-to-speech models, Gemini 3.8 Flash TTS and Gemini 3.8 Flash-Lite TTS, expanding its Gemini family with tools for designing, remixing, and cloning voices. Flash TTS targets creative uses like games, audiobooks, and podcasts, while Flash-Lite TTS serves as a cheaper option for dubbing and voice agents. Both models are available now in the Gemini API and Google AI Studio. The new systems support over 100 languages, let users describe voices in plain text, and access more than 2,000 preset options. Flash TTS can clone a voice from a 30-second sample, requiring verbal consent and adding SynthID watermarks and C2PA credentials to outputs. Google reports a 71.4 score on Hume AI's Voice Design Benchmark, with partners including Figma, HeyGen, and Wondercraft integrating the models.