AssemblyAI's Universal 3.6 Pro tops STT benchmark with 1.77% WER and 91ms median latency

AssemblyAI · x · 2026-09-30

AssemblyAI launched Universal 3.6 Pro, which posts the lowest word error rate (1.77% on 1,000 Pipecat clips, down from 1.93%) among 16 models in Cekura's speech-to-text benchmark. Final-text latency halved: median 91ms (from 180ms) and P95 232ms (from 692ms), beating Google Chirp 3 and others on accuracy and speed.

Related event: AssemblyAI's Universal 3.6 Pro Tops Realtime Speech Benchmarks(3 posts)→

Original post →

More from Models

Models channel →