Speak details GPT-Live-1, OpenAI's new always-listening voice model class
pbbakkum · x · 2026-09-11
Speak engineers spent months working with OpenAI on GPT-Live-1, an entirely new class of voice model that continuously listens and speaks at the same time. They rebuilt their tutoring harness around it and ran internal benchmarks, publishing what they learned about how GPT-Live finally solves turn-taking in voice AI.
Related event: OpenAI launches full-duplex speech model GPT-Live-1 on API(26 posts)→
More from Multimodal
- Transformers v5.17.0 ships major audio additions: TTS, ASR, translation — realmrfakename · 2026-09-11
- Midjourney 8.2 Test Image Surfaces in Community Tease — sebkrier · 2026-09-11
- Astra's first real attempt at generating a classic stickfight animation — repligate · 2026-09-11
- Redditor uses AI to cast himself in a Godzilla-style Kaiju episode — JBOOGZEE · 2026-09-11
- Qwen3-TTS 1.7B hits 1.6x real-time voice cloning on CPU via llama.cpp — alexcovo_eth · 2026-09-11
- Artist synthesizes camera-driven optical flow, quickly hits uncanny territory — pixlpa · 2026-09-11