Google Launches Gemini 3.8 Flash TTS Models Turning Voice Into a Creative Studio
dioscuri · x · 2026-09-24
Google AI Studio introduced two new Gemini 3.8 text-to-speech models, transforming voice generation from static presets into a dynamic creative studio for creators and developers.
Related event: Google Launches Gemini 3.8 Flash TTS with 30-Second Voice Cloning(27 posts)→
More from Multimodal
- FineVision, the 17M-image open VLM dataset from 200+ sources, accepted to NeurIPS — andimarafioti · 2026-09-25
- MiniMax H3 seems overtrained on smiles: 'bored caterpillar' video prompt keeps breaking immersion — episodex86 · 2026-09-25
- Pose Blueprint: A Browser-Based 3D Pose Editor for ComfyUI and ControlNet — OkConfusion6667 · 2026-09-25
- Reddit User Explores AI Art With Only Steps, CFG and Denoise Tweaks, No LoRAs — Extreme_Nice · 2026-09-25
- A sub-$20 LoRA makes Qwen-Image 2.1 rotate transparent objects with a prompt — ben_burtenshaw · 2026-09-25
- Lingbot World v2 runs at 60 FPS, hinting world models could reshape game dev — bingxu_ · 2026-09-25