Gemini 3.8 Flash TTS tops rankings, clones voices from 30 seconds of audio
altryne · x · 2026-09-29
Google is back in voice: Gemini 3.8 Flash TTS is now ranked #1 and can clone a voice from just 30 seconds of audio, with early demos (like a meditation-guide voice) drawing praise. The poster's next live show airs Thu Oct 1, 11am PT from Moscone in SF at CoreWeave #FullyConnected26.
More from Multimodal
- QuiverAI's Arrow 2 Telos hits 1624 Elo, first model to break 1600 on SVG Arena leaderboard — stuffyokodraws · 2026-09-29
- Training FLUX.1 LoRAs on an 8GB RTX 5060: what optimizations work? — Wide_Director_8897 · 2026-09-29
- Flatbed debuts: an AI-native video editor where every asset is individually promptable — rchardkovacs · 2026-09-29
- One Image to a Walkable World: Hyper3D + GPT-6 Rebuilds Scenes in Three.js — Scobleizer · 2026-09-29
- Light Field Primitives: differentiable primitives replace dense ray databases for real-time novel view synthesis — zhenjun_zhao · 2026-09-29
- Face/Head Swap Workflow v2.0 released with full video tutorial — feroszy · 2026-09-29