NetEase Youdao Open-Sources Confucius4-R2T2 Streaming ASR Model

On September 17, NetEase Youdao open-sourced the streaming speech recognition model Confucius4-R2T2 (Real Real-Time Transcription) on GitHub. Built on Qwen3-ASR, it targets true low-latency, high-accuracy streaming recognition and offers an append-only, no-rewriting solution to the state-pollution problem in production voice agent environments. Several bloggers see this as marking streaming ASR's shift from demo to infrastructure.

Confirmed

Why it matters

2026-09-17 ~ 2026-09-17 · 7 related posts

Primary sources