Scale AI Releases Muse Voice Transcribe: SOTA Real-Time Speech-to-Text Model
rohanpaul_ai · x · 2026-09-02
Scale AI has released Muse Voice Transcribe, its first real-time audio perception model. The model achieves state-of-the-art performance in streaming speech-to-text and natively handles speaker diarization and endpointing within a single model.
Related event: Meta Launches Muse Voice Transcribe, Claiming SOTA Streaming ASR(10 posts)→
More from Models
- Heaviside-1: EM Foundation Model 100kx Faster Than Solvers — garrytan · 2026-09-02
- Fable 5.1: Matches High-End Rivals at Lower Cost — daniel_mac8 · 2026-09-02
- User finds regular Claude web sessions appear to run in a Linux VM — majidmanzarpour · 2026-09-02
- Perplexity uses Fable 5.1 as orchestrator with GPT 5.6 as cost-efficient subagents — AravSrinivas · 2026-09-02
- Fable 5.1 takes 1st on Artificial Analysis with a score of 66 — Anxious-Yoghurt-9207 · 2026-09-02
- Perplexity adds Claude Fable 5.1, cutting costs by 37% — perplexity_ai · 2026-09-02