ByteDance Launches Seedance 2.5: Generates 30s Videos, Accepts 50+ Multimodal Inputs
赛博禅心 · wechat · 2026-07-31
ByteDance's Seed team has officially released Seedance 2.5, its latest AI video generation model. The upgrade focuses on long-form storytelling and multimodal capabilities, extending single-shot high-quality video generation to 30 seconds with support for multi-round extensions.
Key upgrades include:
- Extensive Multimodal References: Accepts up to 30 images, 10 videos, and 10 audio clips simultaneously to accurately reproduce complex group scenes and multi-camera setups.
- Precise Timestamp Editing: Allows targeted modifications of characters, actions, and audio within specific timeframes, featuring advanced green screen replacement and camera movement adjustments.
- Visual & Physics Refinement: Systematically reduces the unnatural "AI gloss" and improves the realism of lighting and physical rules.
The model is now available on Jimeng AI and Doubao Pro, with API access coming to Volcano Engine soon. It can also generate synthetic data for industrial manufacturing, embodied AI, and autonomous driving training.
Related event: ByteDance Releases Seedance 2.5: Native 30s Video and Multimodal Control(60 posts)→
More from Multimodal
- MiniMax H3 motion graphics demo impresses, $50K challenge with Picsart opens — egeberkina · 2026-09-16
- After Suno dropped his chords, this user built a MIDI-to-.abc converter so YuE2 preserves them — Saren-WTAKO · 2026-09-16
- Hugging Face team ships 'Building VLMs' book: hands-on guide to vision-language models — andimarafioti · 2026-09-16
- Hugging Face publishes illustrated guide to the 3D generation ecosystem — unofficialmerve · 2026-09-16
- Creator shares 4 imaginary tokens for Midjourney v8.2 to generate metaphorical imagery — LudovicCreator · 2026-09-16
- Zero-cost video generation with a copy-paste Grok prompt that just works — songguoxiansen · 2026-09-16