Google's Gemini 3.8 Flash TTS Lets You Design AI Voices From Text
The Decoder · rss · 2026-09-24
Google is introducing two new text-to-speech models, Gemini 3.8 Flash TTS and Flash-Lite TTS, supporting more than 100 languages. Flash TTS can create new voices from text descriptions, and both models let users add stage directions to individual lines and generate two-voice dialogue from a single script. A voice cloning feature can build a voice profile from a 30-second sample.
More from Models
- Theory: model 'nerfing' may come from mixed heterogeneous inference hardware, not intent — michellechen · 2026-09-24
- GPT-6 Astra (max) reportedly hits 51% pass@5 on ZeroBench, far above 30% SOTA — scaling01 · 2026-09-24
- OpenAI releases MentalHealthBench, an open benchmark built with 80+ clinicians — coherence · 2026-09-24
- "Please Don't Start This with LLMs": Backlash Against Max Prime-Style Model Naming — scaling01 · 2026-09-24
- Limite 1B 'Violetto': tiny open-source model claims competition-math wins over far larger systems — tensorqt · 2026-09-24
- MentalHealthBench: Frontier Models Improving but Gaps Remain in Context-Seeking — thekaransinghal · 2026-09-24