ByteDance Releases SeedAudio1.0
字节跳动Seed · wechat · 2026-07-20
ByteDance Seed has released SeedAudio1.0, an audio creation model designed to generate complete soundscapes rather than isolated audio elements.
The model jointly models vocals, sound effects, and ambient sounds within a unified framework. It supports text + reference audio inputs, 100ms-level time control, generates roughly 2 minutes of audio per run with extension capabilities, and produces natural audio in 20+ languages. The official release also included evaluations across film, podcasts, and livestream e-commerce scenarios, claiming a usability rate of over 90% for most cases.
Related event: ByteDance Releases SeedAudio 1.0 and Opens API(5 posts)→
More from Multimodal
- Non-coder builds full-featured Android ComfyUI client with ChatGPT, submits to Google Play — ComfierUI · 2026-09-11
- FastH3-Live hits 22fps: acceleration node benchmarks and the --vram-headroom trick — spartong945 · 2026-09-11
- Midjourney style code share: --sref 2912175708 — tisch_eins · 2026-09-11
- Astra storyboards plus Minimax H3 per-shot generation boost video success rates — Hailuo_AI · 2026-09-11
- MiniMax H3 MAX nails cooking anime clips: 15-second curry demo with prompts shared — Hailuo_AI · 2026-09-11
- MiniMax Music Production Toolkit 2.5 for ComfyUI adds full mastering chain — Vivid_Promise1700 · 2026-09-11