Alibaba's Qwen-Audio-3.1 ships five voice models, cuts ASR price 95%

千问大模型 · wechat · 2026-09-23

Alibaba's Qwen team released the Qwen-Audio-3.1 family: five models spanning understanding, generation, interaction, and creation — upgraded ASR, TTS, and Realtime, plus new ASR-Next (audio understanding) and TTS-Next (full audio creation) models. Prices dropped across the board: TTS 70%, Realtime 85%, ASR 95%.

Highlights

APIs are live on the Qwen platform and being integrated into agents like Qoder and AI glasses hardware.

Related event: Qwen Launches Five Audio Models with up to 95% Price Cut(4 posts)→

Original post →

More from Multimodal

Multimodal channel →