Pipecat v1.9 adds Meta's Muse Voice Transcribe, the lowest semantic-WER STT model tested
solyarisoftware · x · 2026-09-12
The voice agent framework Pipecat shipped v1.9.0 with support for Meta's new streaming speech-to-text model, Muse Voice Transcribe.
The team maintains Pipecat STT benchmark, an open source test suite that measures STT performance inside a voice agent pipeline. It computes a "semantic word error rate" over 1,000 speech fragments — ignoring transcription differences that don't affect an LLM's understanding of user speech — which they find a better proxy for real accuracy than standard WER algorithms. Muse scored the best (lowest) semantic WER of any model they've tested.
They also measure latency as "time to final segment" of the transcription, a critical metric for voice agents.
More from coding & agent
- Codex quality fixes shipped, reset rolling out to ChatGPT Work users — CtrlAltDwayne · 2026-09-12
- Slack agents start telling each other to back off — paul_cal · 2026-09-12
- Understand-Anything hits 82k GitHub stars turning codebases into explorable knowledge graphs — tom_doerr · 2026-09-12
- dSebastien consolidates free Obsidian guides, 300-tool database and open-source AI skills — dSebastien · 2026-09-12
- Hutter Prize opens an agent swarm challenge to hunt for the ultimate compression algorithm — mervenoyann · 2026-09-12
- Ponytail, the 136k-star repo that makes AI agents write 54% less code — 4310sy · 2026-09-12