Gemini 3.8 Flash TTS Tops Charts, Clones Voices From 30 Seconds of Audio
altryne · x · 2026-09-29
Google is back in voice: Gemini 3.8 Flash TTS has hit #1 on TTS leaderboards and can clone a voice from just 30 seconds of audio — the author tested it as a meditation-guide voice with impressive results.
The linked ThursdAI episode also covers Anthropic and OpenAI shipping cheaper models 101 minutes apart: Claude Opus 5.5 (40% cheaper than Opus 5, cached reads cut to 20¢) and GPT-6 Sol/Luna (half of GPT-5.6's price), plus Meta putting Muse at the center of everything and an open-source Jev clone built in a week.
More from Multimodal
- PrunaAI's P-Video-2 Pro models tie for #2 on Design Arena image-to-video leaderboard at Elo 1325 — guennemann · 2026-09-29
- QuiverAI's Arrow 2 Telos hits 1624 Elo, first model to break 1600 on SVG Arena leaderboard — stuffyokodraws · 2026-09-29
- Training FLUX.1 LoRAs on an 8GB RTX 5060: what optimizations work? — Wide_Director_8897 · 2026-09-29
- Flatbed debuts: an AI-native video editor where every asset is individually promptable — rchardkovacs · 2026-09-29
- One Image to a Walkable World: Hyper3D + GPT-6 Rebuilds Scenes in Three.js — Scobleizer · 2026-09-29
- Light Field Primitives: differentiable primitives replace dense ray databases for real-time novel view synthesis — zhenjun_zhao · 2026-09-29