Meta Releases Muse Voice Transcribe: Real-Time ASR with 20+ Speaker Diarization
AIatMeta · x · 2026-09-02
Meta has released Muse Voice Transcribe, the first real-time audio perception model from Meta Superintelligence Labs. It features streaming ASR with ultra-low latency, diarization supporting over 20 speakers simultaneously, and native multilingual support for 25+ languages with seamless code-switching. The model also leverages keyword and context biasing for improved accuracy and is available via Meta Model API, Meta AI for Mac, and Muse Code.
Related event: Meta Launches Muse Voice Transcribe, Claiming SOTA Streaming ASR(10 posts)→
More from Models
- Anthropic Releases Claude Fable 5.1 and Resets Usage Limits — thesaraharminta · 2026-09-02
- Anthropic Launches Claude 5.1 Series: Focused on Coding and Knowledge Work — aziz4ai · 2026-09-02
- METR reportedly used Redwood's conceptual reasoning benchmark to eval Mythos 5.1 — dfrsrchtwts · 2026-09-02
- Claude Fable 5.1 Launches: 25% Cheaper, Zero Data Retention — prasenx · 2026-09-02
- Fable 5.1 classifiers improved, fewer fallbacks — adonis_singh · 2026-09-02
- Fable 5.1 now testable on Arena in Battle Mode and Agent Mode — arena · 2026-09-02