Alibaba Qwen launches Qwen-Audio-3.0-TTS for real-time and premium speech
千问大模型 · wechat · 2026-07-20
Alibaba Qwen has officially launched Qwen-Audio-3.0-TTS, a new text-to-speech model family with two variants: Flash for real-time interaction and Plus for higher-quality generation.
Key upgrades
- Fine-grained tag control: supports structured tags such as [gasp], [giggles], and [angry] to control breathing, emotion, and delivery details.
- Free-style natural language control: users can describe role, mood, scene, and speaking rate in plain language without speech-specific annotation knowledge.
- Multilingual and dialect coverage: supports 16 languages including Chinese, English, Japanese, Korean, and German, plus new coverage for Arabic, Vietnamese, Malay, and Filipino.
- Dialect synthesis: now covers 20 Chinese dialects, including Cantonese, Chongqing, Northeastern, Shaanxi, Shanghai, Sichuan, and Yunnan.
- Robustness: improved handling of noisy or reverberant reference audio for voice cloning.
Reported results
- On ArtificialAnalysis, Qwen-Audio-3.0-TTS-Plus took the top spot.
- On 16-language evaluations, the family reportedly reached SOTA in 10 languages for word/character error rate, with especially strong gains in Korean and Malay.
- In speaker similarity tests, Plus was best across all 16 languages, while Flash placed second in 15 cases.
The post also says the models are now available on Alibaba Cloud Bailian and via the project blog.
Related event: Alibaba Releases Qwen-Audio-3.0-TTS Voice Synthesis Model(5 posts)→
More from Models
- Same Echo Maze prompt, three frontier models: all passed visually but shipped the same hidden bug — eyishazyer · 2026-09-11
- Benchmark scores drop from 89% to 19% on new evals — how benchmaxxing breaks leaderboard trust — airesearch12 · 2026-09-11
- ChatGPT tells user their question is too hard and to 'accept dumber answers' — phido3000 · 2026-09-11
- Claude is no longer available for minors as Anthropic rolls out age assurance — Muhammad523 · 2026-09-11
- Developer Building a Unified Leaderboard of All Model Benchmark Scores — airesearch12 · 2026-09-11
- Rumor claims Kimi faked performance by serving Claude; DeepSeek new model surprises in evals — realsohamparekh · 2026-09-11