Kyutai releases Voice of Reason, a speech-native reasoning model hitting 77.1% on GSM8K
alexcovo_eth · x · 2026-09-22
Kyutai Labs released Voice of Reason, a speech-native model that reasons and answers entirely in speech: give it a spoken problem with no transcription and no text LLM in the loop. On GSM8K it scores 77.1%, up from 27.3% for GLM-4-Voice.
More from Multimodal
- Reddit Asks for Real-World LongCat-Video Inference Times on RTX 4090 to H100 — Clean_Extreme_3970 · 2026-09-22
- BUPT study: RoPE attention decay causes video diffusion models to violate physics — BUPT-CIST · 2026-09-22
- Tencent ARC's WorldCrafter adds implicit 3D-aware memory to video world models — TencentARC · 2026-09-22
- Grok 4.7 made this in Blender — demo shows the model driving 3D software — iamfakhrealam · 2026-09-22
- Qwen-Image local on a 24GB MacBook Pro takes 5-6 minutes per image — vista8 · 2026-09-22
- Tencent Hunyuan ships Hy Image3.5 preview: +30% human eval win rate at $0.024 per image — TencentHunyuan · 2026-09-22