Google launches Gemini 3.8 Flash TTS, its most expressive audio generation models yet
goyalshaliniuk · x · 2026-09-24
Google introduced Gemini 3.8 Flash TTS and Gemini 3.8 Flash-Lite TTS, described as its most expressive audio generation models yet, aiming to make AI-generated speech sound more natural and engaging. Target use cases span conversational AI, voice assistants, content creation, and interactive experiences. Both are available now via the Gemini API and Google AI Studio.
Related event: Google Launches Gemini 3.8 Flash TTS with 30-Second Voice Cloning(27 posts)→
More from Multimodal
- FineVision, the 17M-image open VLM dataset from 200+ sources, accepted to NeurIPS — andimarafioti · 2026-09-25
- MiniMax H3 seems overtrained on smiles: 'bored caterpillar' video prompt keeps breaking immersion — episodex86 · 2026-09-25
- Pose Blueprint: A Browser-Based 3D Pose Editor for ComfyUI and ControlNet — OkConfusion6667 · 2026-09-25
- Reddit User Explores AI Art With Only Steps, CFG and Denoise Tweaks, No LoRAs — Extreme_Nice · 2026-09-25
- A sub-$20 LoRA makes Qwen-Image 2.1 rotate transparent objects with a prompt — ben_burtenshaw · 2026-09-25
- Lingbot World v2 runs at 60 FPS, hinting world models could reshape game dev — bingxu_ · 2026-09-25