Bland Speech v3 Launches: Beats ElevenLabs & OpenAI, Restores Stroke Patient's Voice
LinusEkenstam · x · 2026-08-05
Bland has launched Speech v3, introduced as the world's first Human Speech Engine. The model was trained on over 100 million real human conversations.
In Design Arena’s Audio Realism benchmark, Speech v3 ranks at the top, outperforming major models including ElevenLabs, Grok, Cartesia, and OpenAI. To demonstrate its capabilities, the team used just 5 seconds of old footage to help James, a 49-year-old father who recently suffered a stroke, successfully regain his voice.
Related event: Bland AI Launches Speech v3, Topping Audio Blind Tests(4 posts)→
More from Multimodal
- Qwen Image 2.1 vs Krea 2 Turbo: stress tests with kilogram-based physique prompts and dense crowd scenes — Crazy-Repeat-2006 · 2026-09-20
- Qwen-Image-2.1 INT8 ConvRot quantized version released on Hugging Face — Easy-Bike1524 · 2026-09-20
- Qwen-Image-2.1 hands-on: new open-source editing standard with 10 reference images and native RGBA output — listopalafoto · 2026-09-20
- Seedance 2.5 clip "Hey Janice, Please Don't Eat Me" wows with absurd drama — TomLikesRobots · 2026-09-20
- UK startup Unit1 raises $20m to stage gigs with hyper-realistic digital avatars — nordicinst · 2026-09-20
- Qwen-Image-2.1 natively generates and edits transparent RGBA images — Alibaba_Qwen · 2026-09-20