AssemblyAI's Universal 3.6 Pro cuts voice-agent transcription errors 45%, adds 14 languages

AssemblyAI · x · 2026-10-06

AssemblyAI's new speech-to-text model, Universal 3.6 Pro, is now live on Vapi. Since most voice agent failures start at transcription, 3.6 Pro makes 45% fewer wrong yes/no confirmations, cuts background speech in transcripts by 30%, and holds the turn until it captures full phone numbers, codes, or emails. It also adds 14 new languages with mid-sentence code-switching; developers can switch agents over in the dashboard.

Original post →

More from Multimodal

Multimodal channel →