Grok Voice Think Fast 2.0 Released: Major Leaps in Voice Reasoning and Noise Robustness
XFreeze · x · 2026-07-30
SpaceXAI launched Grok Voice Think Fast 2.0, its most capable speech-to-speech model to date. The model boasts impressive benchmark numbers:
- Achieves an 82.9% overall Speech-to-Speech Quality Index, 97.2% on speech reasoning, and 56.5% on agentic voice tasks.
- Latency to first audio is just 0.70 seconds, using only 0.4x the reasoning tokens of version 1.0.
It shows significant breakthroughs in real-world transcription. Across thousands of phrases in 24 languages, it delivers 1.5–2x better accuracy than Deepgram Nova 3 and ElevenLabs Scribe v2. In noisy environments or compressed phone calls, this advantage grows to roughly 10x.
Furthermore, the model can reason while speaking, allowing tool calls to begin before the first sentence finishes without adding latency. Slated to become the default grok-voice-latest model on August 5, it is priced at $0.08 per minute. SpaceXAI has already deployed it on Starlink's customer lines, significantly boosting sales conversion and support containment rates.
More from Models
- Researchers Find Anomalous Narrative Fulfillment Tendencies in Claude Opus 5 Base Mode — repligate · 2026-07-30
- Specific Prompt Triggers Anomalous User-Completion Behavior in Claude Opus 5 — matthen2 · 2026-07-30
- New Technologies Like MLA and GRPO are Decentralizing AI Open Source — zephyr_z9 · 2026-07-30
- Users Complain About Claude Opus Being Too Preachy for a Personal Assistant — sidjustice_ · 2026-07-30
- Users Excited as Opus 5 is 'Cooking' with Impressive Capabilities — chrisfirst · 2026-07-30
- Claude Experiences Major Service Outage, Disrupting Access — Polymarket · 2026-07-30