Eleven v4 Tops New TTS Pronunciation Robustness Benchmark, Scoring Above 93% in Three Categories

ArtificialAnlys · x · 2026-09-28

Artificial Analysis introduced Pronunciation Robustness, a benchmark measuring whether TTS models correctly pronounce challenging text across four categories, with human reviewers judging each clip against pre-agreed accepted pronunciations.

Eleven v4 is the only model to score above 93% in three of the four categories.

Original post →

More from Multimodal

Multimodal channel →