Google's Gemini 3.8 Live Speech-to-Speech Models Top Voice Rankings at $0.84/Hour Input
DeepLearningAI · x · 2026-10-07
DeepLearning.AI breaks down Google's new Gemini 3.8 Live speech-to-speech models: single-system listening, reasoning, and responding without relay-style handoffs. The Extended Thinking version ranks first on Artificial Analysis' Speech-to-Speech Index; the standard version ranks second in human-judged blind conversations and costs $0.84 per hour of input audio, the lowest in the index. Both accept image and video input.
More from Models
- Nous Research Launches Hermes Index to Rank Models Inside Hermes Agent — NVIDIAAI · 2026-10-08
- ByteDance Seed Paper Explains Phase Blind Spots in KV Compression Behind DeepSeek's Erratic Long- Context Performance — teortaxesTex · 2026-10-08
- Gary Marcus Slams OpenAI's Vague Math Proof Report: Zero Details, Won't Pass Peer Review — GaryMarcus · 2026-10-08
- Perplexity releases pplx-embed-v2-late: OCR-free late-interaction embeddings topping retrieval benchmarks — perplexity_ai · 2026-10-08
- Anonymous stealth LLM 'Space Bunny Alpha' tops OpenRouter with 22% usage share — maferase · 2026-10-08
- Check Point Breaks Decision Model Jev for About 50 Cents per Attack — evilsocket · 2026-10-08