ByteDance Launches SeedRealtime: A Native Audio-Visual Full-Duplex LLM

iamrobotbear · x · 2026-08-06

ByteDance has launched SeedRealtime, a native audio-visual full-duplex LLM. Using a unified architecture, the model natively fuses audio, video, and text to enable real-time interaction over continuous multimodal streams, delivering a "watch, listen, and speak" experience.

It is now live and free on the Doubao App, supporting real-time voice and video conversations. In complex scenarios like group dinners, the model can accurately map names to faces after a brief introduction and consistently distinguish between multiple speakers during chaotic, overlapping conversations.

Related event: ByteDance Launches Full-Duplex SeedRealtime Model on Doubao(7 posts)→

Original post →

More from Models

Models channel →