Speech-to-Speech Model Analysis: Grok Leads Task Success, Gemini Top Preference
ArtificialAnlys · x · 2026-08-21
Artificial Analysis released a comprehensive comparison of Speech-to-Speech models and providers, evaluating reasoning quality, conversational dynamics, speed, and pricing.
- Overall Preference: Gemini 3.1 Flash Live Preview leads with 1046 Elo at $1.50/hour.
- Task Success Rate: Grok Voice Think Fast 2.0 High achieves 94.7% at $4.80/hour.
- GPT-Realtime: GPT-Realtime-2.1 High follows with a 91.5% task success rate.
The report includes detailed leaderboards, benchmarking methodology, and example conversations.
More from Models
- V4-Flash-Vision tested: 2x faster, cheaper and better than V4-Flash-0731 — cedric_chee · 2026-08-21
- Kimi K3.1 quietly testing on Code Arena; Ox Alpha said to be Zhipu's GLM 5.3 Flash — i_dg23 · 2026-08-21
- OpenRouter's stealth model Ox Alpha sparks guessing game over its identity — AccBalanced · 2026-08-21
- ThursdAI weekly: Qwen 27B and GLM 5.3 beat GPTs; OpenAI pauses RL for security — thursdai_pod · 2026-08-21
- Ox Alpha's stunning fluid sim 'one-shot' appears to be a copy of an existing GitHub repo — scaling01 · 2026-08-21
- Creator's GPT Image 3 wishlist: perfect pixel art, zero noise, working spritesheets — Angaisb_ · 2026-08-21