Meta Superintelligence Labs Launches Muse Voice Transcribe, First Real-Time Audio Perception Model
aigclink · x · 2026-09-02
Meta Superintelligence Labs has introduced Muse Voice Transcribe, described as its first real-time audio perception model. Key capabilities:
- Real-time streaming ASR with endpointing;
- Diarization supporting 20+ speakers;
- 25+ languages with seamless code-switching;
- Improved accuracy via language, keyword and context biasing.
Meta says the model ranks first on Artificial Analysis streaming speech-to-text and on public diarization benchmarks (as of September 1, 2026). A live microphone transcription demo is available; audio is not stored.
More from Models
- Fable 5.1 review: Tends to act as a 'manager' and plan globally — AlchainHust · 2026-09-02
- Elon Musk announces Grok 4.7 release in 10 days — XFreeze · 2026-09-02
- Speculation: Google merging world model and diffusion into Gemini 4 — haider1 · 2026-09-02
- Fable 5.1 system prompt leak reveals search mechanics of the "most powerful" AI — gaganghotra_ · 2026-09-02
- Claude one-shot identifies a person from raw weights at ~75% accuracy — repligate · 2026-09-02
- Fable 5 vs 5.1 self-portrait comparison re-run without the new system prompt — repligate · 2026-09-02