Qwen-Audio-3.0-TTS launches with 16 languages and better voice cloning
airesearch12 · x · 2026-07-21
Qwen releases Qwen-Audio-3.0-TTS in two variants:
- Flash for real-time interaction
- Plus for higher-quality generation
Notable additions include support for 16 languages, natural-language style control, fine-grained tags for non-verbal details, and more robust voice cloning from imperfect audio.
Related event: Alibaba Releases Qwen-Audio-3.0-TTS Voice Synthesis Model(5 posts)→
More from Multimodal
- Midjourney style code share: --sref 2912175708 — tisch_eins · 2026-09-11
- Astra storyboards plus Minimax H3 per-shot generation boost video success rates — Hailuo_AI · 2026-09-11
- MiniMax H3 MAX nails cooking anime clips: 15-second curry demo with prompts shared — Hailuo_AI · 2026-09-11
- MiniMax Music Production Toolkit 2.5 for ComfyUI adds full mastering chain — Vivid_Promise1700 · 2026-09-11
- New Node Finder for ComfyUI ranks fresh nodes by star velocity and recency — Luke2642 · 2026-09-11
- Using a finisher move on one mosquito with MiniMax H3 MAX — the bug survives — Hailuo_AI · 2026-09-11