MOSS Speech Transcription Model Goes Viral

OpenMOSS-Team · hf · 2026-07-09

OpenMOSS-Team's MOSS-Transcribe-Diarize is gaining traction on Hugging Face as an audio-to-text model pipeline for speech recognition and speaker diarization.

Tags indicate support for ASR, diarization, and timestamp-asr, highlighting its focus on transcription, segmentation, and timestamp alignment.

Related event: MOSS Releases Open-Source 0.9B Audio Transcription Model Supported by vLLM(5 posts)→

Original post →

More from Multimodal

Multimodal channel →