Gemini 3.8 Flash TTS tops Artificial Analysis pronunciation benchmark at 89.5%
ArtificialAnlys · x · 2026-09-24
Artificial Analysis released updated results for its Pronunciation Robustness benchmark: Gemini 3.8 Flash TTS takes #1 at 89.5%, ahead of Gemini 3.1 Flash TTS (88.2%), SpaceXAI TTS (87.6%) and Gemini 3.8 Flash-Lite TTS (87.4%). By category, Gemini 3.8 Flash TTS leads on Contextually Appropriate pronunciation (97.9%) and Expanding Shorthand (86.1%), SpaceXAI TTS leads on Preserving Exact Sequences (85.7%), and Qwen-Audio-3.0-TTS-Plus tops Standalone Terms (95.5%). Listenable samples are included.
More from Multimodal
- Claude Opus 5.5 generates a music video in ~1 shot: "not good, but not without interest" — NathanpmYoung · 2026-09-24
- Gemini 3.8 Flash TTS and Flash-Lite TTS land on Merge Gateway, top Hume voice quality index — shensi · 2026-09-24
- LemonSlice Launches Character World Model-1, a Real-Time Interactive Avatar Model — mhdfaran · 2026-09-24
- Reverse workflow: unpack a reference image with Extract Prompt, then remix it — JaynitMakwana · 2026-09-24
- Unverified claim: 'GPT-6 Astra' builds full video projects via Codex + Dreamina CLI — JaynitMakwana · 2026-09-24
- MiniMax H3 roundup: video VAE 2.2x faster encoding, music model under 12GB — optimisticalish · 2026-09-24