How to clone your own voice with Gemini 3.8 TTS: Phil Schmid's hands-on guide
_philschmid · x · 2026-09-24
Phil Schmid published a hands-on guide for the newly available Gemini 3.8 Flash TTS and Flash-Lite TTS, which top Hume's Voice Design Benchmark and Voice Arena in 6 languages. Key tips: record two clips with the same mic and room at 24kHz mono, use -f alsa -i default on Linux (re-run after granting macOS mic permission), and read one of the 25 exact consent sentences. The guide shows both voice replication and creating a new voice from a single sentence, with a prompt you can hand to your agent.
More from Multimodal
- User generates a music video from old material with Opus 5.5 — repligate · 2026-09-25
- Same prompt, Opus 5.5 one-shot video generation put to a public retest with different tools — drrickio · 2026-09-25
- One prompt: Claude agent wired to Runway MCP delivers a Netflix-style superintelligence doc — CurieuxExplorer · 2026-09-25
- Seedance 2.5 + GPT Image 2.5 One-Shot the Most Iconic Sci-Fi Rivalry — CurieuxExplorer · 2026-09-25
- Opus 5.5 writes songs and music videos entirely from code in a playable demo — pbaylies · 2026-09-25
- Study: Reasoning hurts 15.7% of multimodal embeddings; training-free SURE router fixes it — _reachsumit · 2026-09-25