Google releases Gemini 3.8 Live and 3.5 Transcribe audio models for real-time voice apps
gaganghotra_ · x · 2026-09-16
Google AI Studio announced new Gemini audio models — Gemini 3.8 Live and 3.5 Transcribe — now available via the Gemini API and Google AI Studio, enabling developers to build real-time conversational voice applications.
More from Multimodal
- Keeping One Man and His Dog Consistent Across an Entire AI Short Film with Kling 3.0 — Loretaro · 2026-09-16
- MiniMax H3 lands on Together AI: 33B omni-modal model generates 2K clips with native stereo audio — togethercompute · 2026-09-16
- Indie dev replaces flat ghosts with rigged, animated MeshyAI 3D monsters — abandonedmuse · 2026-09-16
- Shareable ChatGPT Image Prompt Keeps Facial Identity While Placing You in a Paris Café Scene — RachelVT42 · 2026-09-16
- FastVideo FastH3 V2 open weights drop with ComfyUI workflow and 8-step version — fruesome · 2026-09-16
- fal's H3 Max: RL post-training cuts 5s video generation from 120s to under 3s — isidentical · 2026-09-16