VAmoS Pro voice-agent benchmark: Grok leads tasks, GPT-Live fastest, Gemini most noise-robust
davlanade · x · 2026-10-02
VAmoS Pro, billed as the most realistic voice-agent benchmark yet, is out with early results showing each lab winning a different axis: xAI's Grok Voice Think Fast 2.0 leads on task completion, OpenAI's GPT-Live 1 has the fastest response time, and Google's Gemini 3.8 Live is the most robust to noise. Details are in the linked thread—useful reference for anyone picking a voice model for agents.
More from Models
- Compound AI adds Sol 6.1 and argues products should abstract model choice away from users — peterjliu · 2026-10-02
- AI keeps getting cheaper while plans get pricier: the pricing paradox — haider1 · 2026-10-02
- Five AI models try Hemingway's style — GPT, Gemini and Claude end up closer to parody — rvp · 2026-10-02
- Qwen 4 reportedly matches Fable 5 in early samples; 27B due early October — petrusenko_max · 2026-10-02
- ChatGPT Business user finds the seat without custom instructions works far better — BraveBrush8890 · 2026-10-02
- GPT-6.1 Sol tops Animation Bench, first model to cross 0.5 on motion consistency — himanshustwts · 2026-10-02