Google launches Gemini 3.8 Flash TTS and Flash-Lite TTS with prompt-based voice design
Google DeepMind
Google released two text-to-speech models: Gemini 3.8 Flash TTS and Flash-Lite TTS, available in the Gemini API and AI Studio. They support custom voice design from text descriptions, voice cloning from about 30 seconds of audio, multi-speaker dialogue, 2000+ voices and 100+ languages.
Why it matters
Benchmark-topping speech generation with voice cloning built into a mainstream frontier API.
Importance: 3/5
Notable audio release from a frontier lab, official+media