DeepMind ships Gemini 3.8 Flash TTS with native multi-speaker overlap and laughter
kastnerkyle · x · 2026-09-24
Google DeepMind launched two new TTS models: Gemini 3.8 Flash TTS, which can design voices with distinct accents and characteristics, and Flash-Lite TTS, built for efficiency and scale with a production-ready voice library. Team member MosesOh4 highlights the key breakthrough—native multi-speaker speech where models can laugh together, backchannel and overlap naturally—paving the way from TTS and NotebookLM-style dialogs to real-time assistants. More audio generation models are teased.
More from Multimodal
- Prompt Templates: Turn Product Photos Into Pro Ads and Expand Images With ChatGPT — TawohAwa · 2026-09-24
- Prompt share: pastel manga character study with multi-angle sheets — azed_ai · 2026-09-24
- Dreamina's rebuilt web app adds node-branching video workflow, $1.5 promo until Oct 9 — Div_pradeep · 2026-09-24
- Users complain GPT Image 2.5 still has noise artifacts OpenAI promised to fix 6 months ago — AdStreet4350 · 2026-09-24
- Treating AI video as a production pipeline: building a reusable cinematic producer agent in CREAO — FellMentKE · 2026-09-24
- Open-source ComfyUI node restores lost template filters for Local, Partner and Credit workflows — linus74RN · 2026-09-24