Grok Voice 2.0 Tops Tau Voice Benchmark with Massive Leaps in Full Duplex Scores
ArtificialAnlys · x · 2026-07-30
Artificial Analysis released detailed benchmark results for Grok Voice Think Fast 2.0 High. The model performs exceptionally well across multiple evaluations:
- τ-voice Bench (Agentic Performance): Takes the #1 spot at 56.5%, ahead of Qwen Audio 3.0 Realtime Plus (54.6%) and its predecessor (52.1%).
- Big Bench Audio (Speech Reasoning): Achieves 97.2%, slightly better than the previous generation, sitting just behind the leader Qwen Audio 3.0 (99.2%).
- Full Duplex Bench (Conversational Dynamics): Jumps significantly to 95.1% from the previous 77.8%, narrowly trailing Qwen Audio 3.0 (98.4%).
Additionally, the model ranks #2 on the Speech-to-Speech Index (82.9%) and is among the fastest, with a Time to First Audio of just 0.70s.
More from Models
- Researchers Find Anomalous Narrative Fulfillment Tendencies in Claude Opus 5 Base Mode — repligate · 2026-07-30
- Specific Prompt Triggers Anomalous User-Completion Behavior in Claude Opus 5 — matthen2 · 2026-07-30
- New Technologies Like MLA and GRPO are Decentralizing AI Open Source — zephyr_z9 · 2026-07-30
- Users Complain About Claude Opus Being Too Preachy for a Personal Assistant — sidjustice_ · 2026-07-30
- Users Excited as Opus 5 is 'Cooking' with Impressive Capabilities — chrisfirst · 2026-07-30
- Claude Experiences Major Service Outage, Disrupting Access — Polymarket · 2026-07-30