Scale AI Releases Muse Voice: SOTA Streaming Speech-to-Text Model

alexandr_wang · x · 2026-09-02

Scale AI has released Muse Voice Transcribe, its first real-time audio perception model. It achieves SOTA performance in streaming speech-to-text and natively handles speaker diarization and endpointing within a single model.

Related event: Meta Superintelligence Labs Launches Muse Voice Transcribe, Its First Real-Time Audio Perception Model(8 posts)→

Original post →

More from Models

Models channel →