Grok Voice Beats GPT Realtime in VulcanBench with Lower 'Voice Tax'
XFreeze · x · 2026-08-02
In a head-to-head voice model comparison by VulcanBench, xAI's Grok Voice Think Fast 2.0 outperformed OpenAI's GPT Realtime.
Both models were asked the same 200 questions in text and spoken audio formats. The results show:
- Grok: 99.0% text accuracy → 95.7% audio accuracy (-3.3 pp)
- GPT Realtime: 97.5% text → 93.5% audio (-4.0 pp)
Grok not only achieved higher overall accuracy but also had a smaller "voice tax"—the accuracy lost when switching from typed text to spoken audio. Furthermore, Grok completed each turn nearly 2× faster than GPT, although GPT started speaking sooner. This demonstrates that effective voice AI must listen, reason, and solve tasks correctly in real-time, beyond just sounding human.
More from Models
- Google's Astra Model Breaks Through Math, Disproves Connes' Conjecture — npew · 2026-08-02
- Opus 3.5 Personality Shift? User Blasts Model Updates for Killing Diversity — liminal_bardo · 2026-08-02
- Extrapolation: iPhones May Run Claude 3 Opus-Level Agentic Models by 2027 — appenz · 2026-08-02
- Beyond Benchmarks: Speed and Cost Dictate Agent Survival in Production — RachelVT42 · 2026-08-02
- TokenRouter Offers 50M Free Tokens for Kimi K3 API Access — dr_cintas · 2026-08-02
- User Complaints: DeepSeek Ignores Rule Prompts, Falls Behind Qwen in Practice — Juulk9087 · 2026-08-02