ByteDance Releases SeedAudio1.0
字节跳动Seed · wechat · 2026-07-20
ByteDance Seed has released SeedAudio1.0, an audio creation model designed to generate complete soundscapes rather than isolated audio elements.
The model jointly models vocals, sound effects, and ambient sounds within a unified framework. It supports text + reference audio inputs, 100ms-level time control, generates roughly 2 minutes of audio per run with extension capabilities, and produces natural audio in 20+ languages. The official release also included evaluations across film, podcasts, and livestream e-commerce scenarios, claiming a usability rate of over 90% for most cases.
Related event: ByteDance Releases SeedAudio 1.0 and Opens API(5 posts)→
More from Multimodal
- Storyboard-first workflows are making AI dance videos and influencers more consistent — aftahi_ai · 2026-07-22
- Interactive video should be judged by responsiveness, not just frame quality — Soggy_Limit8864 · 2026-07-22
- Runpod MCP and Claude help spin up image and video generation workflows — 802high · 2026-07-22
- Midjourney prompt turns a bee into a glitching pixel explosion — michaelrabone · 2026-07-22
- A physics reward can improve video generation without creating a real physics engine — Dapper-Drawer4546 · 2026-07-22
- HeyGen adds a media-sourcing skill for coding agents with 75k images and 10k tracks — HeyGen · 2026-07-22