Meta releases Muse Voice Transcribe, a real-time audio perception model
bowenc0221 · x · 2026-09-02
Meta Superintelligence Labs introduces Muse Voice Transcribe, their first real-time audio perception model. It delivers streaming ASR, diarization for 20+ speakers, and endpointing. The model features multilingual code-switching and improved accuracy via language, keyword, and context biasing. It ranks first on ArtificialAnalysis's streaming speech-to-text and public diarization benchmarks.
Related event: Meta Launches Muse Voice Transcribe, Claiming SOTA Streaming ASR(10 posts)→
More from Models
- Heaviside-1: EM Foundation Model 100kx Faster Than Solvers — garrytan · 2026-09-02
- Fable 5.1: Matches High-End Rivals at Lower Cost — daniel_mac8 · 2026-09-02
- User finds regular Claude web sessions appear to run in a Linux VM — majidmanzarpour · 2026-09-02
- Perplexity uses Fable 5.1 as orchestrator with GPT 5.6 as cost-efficient subagents — AravSrinivas · 2026-09-02
- Fable 5.1 takes 1st on Artificial Analysis with a score of 66 — Anxious-Yoghurt-9207 · 2026-09-02
- Perplexity adds Claude Fable 5.1, cutting costs by 37% — perplexity_ai · 2026-09-02