Monologue launches in-house dictation model mono-1: 55% fewer edits, 3x faster than API pipeline

every · x · 2026-09-24

Dictation app Monologue introduced mono-1, its in-house dictation model built over six months to replace its API-based pipeline. On WildSpeech-Bench (1,100 English clips covering noise, stuttering, overlapping speech), mono-1 hit a 7.22% WER—second only to OpenAI's GPT-4o Transcribe at 6.68%—with a 109 ms median raw-transcription latency, tied fastest with Cartesia Ink. Versus the old pipeline, mono-1 needs 55% fewer estimated edits and delivers finished dictation 3x faster by median response time, handling corrections, formatting, context, and writing preferences.

Original post →

More from Models

Models channel →