Bland AI Launches Speech v3, Topping Audio Realism Blind Benchmark

EXM7777 · x · 2026-08-05

Bland AI has introduced Bland Speech v3, a new text-to-speech model described by the company as the first "Human Speech Engine."

Trained on over 100 million real human conversations, the model ranks first on the third-party Audio Realism Bench with an Elo score of 1237. It outperforms major competitors including Grok TTS, GPT Realtime 2, Gemini 2.5 Pro TTS, and ElevenLabs. The company also showcased a use case where they used just 5 seconds of old audio footage to help a 49-year-old father who lost his voice to a stroke speak again.

Original post →

More from Multimodal

Multimodal channel →