Google Launches Gemini 3.8 Flash TTS with 30-Second Voice Cloning

On September 23, Google DeepMind released two next-generation text-to-speech models: Gemini 3.8 Flash TTS for high-fidelity creative work (games, audiobooks, podcasts) and Gemini 3.8 Flash-Lite TTS for high-volume, low-cost scenarios. Google calls them its most expressive audio generation models with SOTA performance; they are available in AI Studio and Gemini platforms, with generated audio watermarked.

Confirmed

Unconfirmed

Why it matters

2026-09-22 ~ 2026-09-24 · 27 related posts

Full story(4 episodes)→

Primary sources

18 near-duplicate retellings: AI_Andrew · _philschmid · patloeber · ArtificialAnlys · OfficialLoganK · rseroter · OfficialLoganK · kimmonismus · Arindam_1729 · Recoil42 · kastnerkyle · AI_Andrew · giffmana · petrusenko_max · altryne · dioscuri · rseroter · goyalshaliniuk