GPT-Live-1 Meets Your Own Backend: Turn-Taking, Silence-Filling, and Who Owns the Call
Slight_Republic_4242 · reddit · 2026-09-15
Dograh's founder breaks down wiring GPT-Live-1 as a voice layer with any backend: the benchmark score was measured with GPT-6 Astra behind it, so your number depends on your own stack. Two hard-won lessons: (1) the voice and reasoning models don't share a clock — GPT-Live-1 is trained to never sit in silence, and on a customer call, filling is committing; (2) you hand turn-taking to the vendor — OpenAI owns the VAD signals, Gemini Live splits it differently, Ultravox hands it back. Speech-to-speech also sacrifices the cascaded stack's observability, and one vendor now owns the whole call while training on your data and serving competitors. 'The key takes an afternoon. The system around the key is the product.'
Related event: GPT-Live-1 Benchmark Scores Depend on Backend Model(2 posts)→
More from coding & agent
- WoW add-on lets you chat with Claude Code or Codex agents while grinding levels — nptacek · 2026-09-16
- Why SwiftUI feels janky: it ditches Core Animation, the backbone of iOS smoothness — dotey · 2026-09-16
- GMI launches MCP server exposing 150+ multimodal models to Claude, ChatGPT and Cursor — _jaydeepkarale · 2026-09-16
- Lovable rewrites Vite dev server in Rust: 2x faster cold starts, 4x less memory; Evan You responds — cnakazawa · 2026-09-16
- Codex CLI Users Angry: Two Unwanted Features Added in Quick Succession, One Wastes Tokens — ssh4net · 2026-09-16
- Google's Gary Illyes: keep important content in HTML as LLMs crawl the web — lilyraynyc · 2026-09-16