Google launches Gemini 3.8 Flash and Flash-Lite TTS with voice cloning and watermarking
petrusenko_max · x · 2026-09-24
Google launched Gemini 3.8 Flash TTS and Flash-Lite TTS on Sept 23, 2026. The models can build custom voices from prompts or replicate existing ones, take line-by-line direction on pacing and emotion, and ship with watermarking. Available in AI Studio, the Gemini API, and Vids.
Related event: Google Launches Gemini 3.8 Flash TTS with 30-Second Voice Cloning(27 posts)→
More from Multimodal
- FineVision, the 17M-image open VLM dataset from 200+ sources, accepted to NeurIPS — andimarafioti · 2026-09-25
- MiniMax H3 seems overtrained on smiles: 'bored caterpillar' video prompt keeps breaking immersion — episodex86 · 2026-09-25
- Pose Blueprint: A Browser-Based 3D Pose Editor for ComfyUI and ControlNet — OkConfusion6667 · 2026-09-25
- Reddit User Explores AI Art With Only Steps, CFG and Denoise Tweaks, No LoRAs — Extreme_Nice · 2026-09-25
- A sub-$20 LoRA makes Qwen-Image 2.1 rotate transparent objects with a prompt — ben_burtenshaw · 2026-09-25
- Lingbot World v2 runs at 60 FPS, hinting world models could reshape game dev — bingxu_ · 2026-09-25