Meta releases Muse Voice Transcribe with top streaming accuracy
ArtificialAnlys · x · 2026-09-02
Meta Superintelligence Labs released Muse Voice Transcribe, the first streaming Speech-to-Text model. It achieves #1 Final Transcript accuracy on AA-WER Streaming with 3.1% WER at 0.16s latency, supporting 70+ languages. Priced at $0.18/hour, it undercuts competitors like Cartesia Ink-2 and ElevenLabs.
Related event: Meta Launches Muse Voice Transcribe, Claiming SOTA Streaming ASR(10 posts)→
More from Multimodal
- Meta Avatars 2.0: Stylized FACS Implementation Details — SergiCaballer · 2026-09-02
- Book 'KI-KUNST' explores the creativity and controversy of AI art — Merzmensch · 2026-09-02
- Interactive brain model: AI traces anatomy when you hear 'pass the salt' — mikeyk · 2026-09-02
- AI-Generated Drink Ad Features Eye Reflections and Splashes — anthara_ai · 2026-09-02
- Prompt for Fabric and Light Title Sequence via MiniMax H3 — umesh_ai · 2026-09-02
- Interest shifts to Meta's real-time voice transcription model — IndraVahan · 2026-09-02