Gemini 3.8 Live launches: #1 voice model at $0.005/min with async background thinking
_philschmid · x · 2026-09-16
Google released Gemini 3.8 Live and 3.8 Live Extended Thinking, continuing its real-time voice agent push after last month's 3.5 Transcribe.
- 82.6 score (#1) on Artificial Analysis Quality Index
- Pricing: $0.005/min input, $0.018/min output
- 35.1 (#1) on agentic task completion (τ-banking)
- Extended Thinking runs background reasoning and asynchronous tool calls while the conversation keeps flowing, coordinating multi-step tool use without breaking the dialogue
Available via the Gemini API in AI Studio, Google Search, and the Gemini App, plus partner plugins for LiveKit, Pipecat, LangChain, and Vercel.
More from Models
- OpenRouter spend flips to OpenAI over Anthropic for first time in 2.5 years — firstadopter · 2026-09-16
- KD in mid-training favors reasoning over factual recall, AI2/UW paper finds; Switch Distillation proposed — LukeZettlemoyer · 2026-09-16
- DoorDash, Siemens, Airbnb shift to cheap Chinese open-weight models — carlbfrey · 2026-09-16
- StepFun launches StepAudio 3: five audio models topping realtime voice leaderboards — StepFun_ai · 2026-09-16
- Abacus.AI says Smaug Flash fixes open-source models' tool-call hangs in production — bindureddy · 2026-09-16
- Voice mode breaking up, Codex erroring: reliability still far from AGI-ready — koltregaskes · 2026-09-16