Google launches Gemini 3.8 Flash TTS with 100+ languages and 2,000+ prebuilt voices
AI_Andrew · x · 2026-09-23
Google released two speech generation models: Gemini 3.8 Flash TTS for expressive, nuanced voice acting and 3.8 Flash-Lite TTS for high-volume, cost-efficient scale, both covering 100+ languages.
Key features:
- Voice Design: craft custom voices across languages, dialects and accents from text prompts
- 2,000+ prebuilt voices: high-fidelity library spanning diverse styles and personas
- Voice replication: fast, authentic cloning with built-in consent verification
- Multi-speaker dialogue: natural pacing between distinct voices
Available now in Google AI Studio for developers and creators.
Related event: Google launches Gemini 3.8 Flash TTS with instant voice cloning(14 posts)→
More from Multimodal
- DIY Minimax H3 workflow adds multi-image + audio + video reference inputs — TheNeonGrid · 2026-09-24
- How PixVerse R2 works: dynamic chunks adapt to input, multi-timescale memory keeps worlds consistent — SarahAnnabels · 2026-09-24
- Looking for a model to label each subtitle line with the speaker's identity — dtdisapointingresult · 2026-09-24
- fal enterprise spend quadruples in six months; compute is generative media's binding constraint — isidentical · 2026-09-24
- APOB pairs with Seedance 2.5 to turn AI influencer vlogs into consistent 30-second cinematic stories — aftahi_ai · 2026-09-24
- Opus 5.5 autonomously produces a film-history short in 90 minutes using 6% of usage limits — GabGarrett · 2026-09-24