←

Google launches Gemini 3.8 Flash TTS and Flash-Lite TTS with prompt-based voice design

Google DeepMind

Audio official + media 3 src. ~1 min

Google released two text-to-speech models: Gemini 3.8 Flash TTS and Flash-Lite TTS, available in the Gemini API and AI Studio. They support custom voice design from text descriptions, voice cloning from about 30 seconds of audio, multi-speaker dialogue, 2000+ voices and 100+ languages.

Why it matters

Benchmark-topping speech generation with voice cloning built into a mainstream frontier API.

Importance: 3/5

Notable audio release from a frontier lab, official+media

Sources