Google launches Gemini 3.8 Live real-time speech models with 97-language support
Google officially launched two real-time audio conversation models on September 16—Gemini 3.8 Live and Gemini 3.8 Live Extended Thinking—calling them its most advanced Gemini audio models to date, achieving SOTA performance at frontier-model pricing and performance. Developers can try them online in Google AI Studio.
Confirmed
- Both models support 97 languages with seamless mid-conversation switching, automatic language detection, and barge-in during voice conversations
- Shared features include upgraded reasoning, near-real-time visual understanding, and background (asynchronous) tool calls that don't interrupt the chat
- Gemini 3.8 Live Extended Thinking adds extended thinking; DeepMind's official demo showed it acting as a coding tutor, reasoning and explaining without breaking the conversation
- According to philschmid, Gemini 3.8 Live tops the Artificial Analysis Quality Index with a score of 82.6
- The day before launch (September 15), testingcatalog cited Bedros Pamboukian's finding that entries for both models had appeared in the GCP console quotas and metrics pages, corroborating the pre-launch leak
Why it matters
- Continuing the real-time voice agent line started by last month's Gemini 3.5 Transcribe, it is seen as a direct answer to GPT Live (per testingcatalog)
- The combined ability to "speak, think, and run background tasks without interrupting the user's flow" marks the evolution of voice assistants from pure conversation toward real-time agents that can execute tasks
2026-09-15 ~ 2026-09-16 · 14 related posts
Primary sources
- Gemini 3.8 Live and Extended Thinking variants spotted on GCP quota page, answering GPT Live 1 — testingcatalog · 2026-09-15
- Gemini 3.8 Flash Live Spotted in Google Cloud Backend, Launch Rumored Within Days — koltregaskes · 2026-09-16
- [source] Google DeepMind launches Gemini 3.8 Live and Live Extended Thinking — GoogleDeepMind · 2026-09-16
- DeepMind demos Gemini 3.8 Live as a coding tutor that narrates its reasoning — GoogleDeepMind · 2026-09-16
- [source] Google launches Gemini 3.8 Live and 3.8 Live Extended Thinking with 97-language real-time voice — GoogleAI · 2026-09-16
- Gemini 3.8 Live launches: #1 voice model at $0.005/min with async background thinking — _philschmid · 2026-09-16
- Gemini 3.8 Live now tryable live in Google AI Studio — _philschmid · 2026-09-16
- Google launches Gemini 3.8 Live, a SOTA live audio model with 97-language switching — OfficialLoganK · 2026-09-16
- Google's new Gemini audio models power Search Live conversations globally — gaganghotra_ · 2026-09-16
- Google Ships Gemini 3.8 Live Real-Time Voice Models with Search Live Troubleshooting — AI_Andrew · 2026-09-16
4 near-duplicate retellings: GoogleAI · _philschmid · OfficialLoganK · minchoi