Artificial Analysis releases full streaming STT leaderboard and methodology

ArtificialAnlys · x · 2026-10-02

Artificial Analysis published full results and methodology for its streaming speech-to-text benchmark. The AA-WER Streaming index uses 8 hours of audio weighted across AA-AgentTalk (50%), VoxPopuli (25%) and Earnings22 (25%), covering diverse accents, domain language and tough acoustics, comparing 37+ models on WER, latency after end of speech and price, with a Pareto-frontier view split by proprietary vs open weights.

Original post →

More from Models

Models channel →