Bland Launches Speech v3 Human Speech Engine, Outperforming ElevenLabs and OpenAI
ycombinator · x · 2026-08-05
Bland has introduced Speech v3, billed as the world's first Human Speech Engine. Trained on over 100 million real human conversations, the model targets businesses, developers, and creators.
Speech v3 ranked first in the Design Arena Audio Realism benchmark, beating out major competitors including ElevenLabs, Grok, Cartesia, and OpenAI. To showcase its capabilities, Bland used just 5 seconds of old footage to help a 49-year-old father who lost his voice to a stroke speak again.
Related event: Bland AI Launches Speech v3, Topping Audio Blind Tests(4 posts)→
More from Multimodal
- AudioSlopServer: host multiple audio diffusion models on one GPU — SteveLittleFish · 2026-09-20
- AI video tool's 'next episode' button auto-generates endless bingeable episodes — Kyrannio · 2026-09-20
- NoSpoon microdrama agent auto-generates posters via Grok Imagine — Kyrannio · 2026-09-20
- MiniMax H3 generates a stunning 'Dragon Cave' video on Reddit — apoke890 · 2026-09-20
- Runway wants to turn AI video generation into a real-time controllable live stream — The Decoder · 2026-09-20
- AI-generated series 'Omniverse 24/7' releases episode 5 — zlausd · 2026-09-20