Unverified: Google's Gemini 3.8 Live brings multi-step reasoning to realtime voice
xiaohu · x · 2026-09-16
A Chinese blogger claims Google released Gemini 3.8 Live and an Extended Thinking variant pushing realtime voice into multi-step reasoning (unverified; the version naming is questionable). Base version targets low-latency, high-concurrency deployments with automatic language switching across 97 languages; Extended Thinking adds synced "thinking while speaking," asynchronous background tool calls without interrupting conversation, and near-realtime visual understanding — e.g., playing chess from the board or turning a sketch into React code over voice chat.
More from Models
- V4.1 Flash called the first open model that feels proto-AGI — teortaxesTex · 2026-09-16
- Claude Opus 5's PRs proactively confess every mistake made while coding — repligate · 2026-09-16
- Six efficiency breakthroughs labs didn't see coming upend semiconductor demand assumptions — bookwormengr · 2026-09-16
- A common jailbreak: third-person role-play gradually blurs lines to widen the model's Overton window — BlancheMinerva · 2026-09-16
- cocktail-peanut predicts Jev will go open weights, seeing more potential locally than as an API — cocktailpeanut · 2026-09-16
- Speculation mounts OpenAI's mysterious 'new model' is a fresh pretrain, not an RL run — teortaxesTex · 2026-09-16