StepFun and ACE Studio launch music foundation model StepAudio 3 Music
StepFun_ai · x · 2026-09-17
StepFun and ACE Studio present StepAudio 3 Music, StepFun's first music generation foundation model, turning a prompt plus lyrics into a complete song.
Users can specify genre, mood, vocal character, instruments, key, BPM, and song structure. One model covers four workflows: song generation, instrumental generation, music covers, and vocal-to-song arrangement. A Hugging Face Space, API docs, and a technical report are available.
Related event: StepFun and ACE Studio Release StepAudio 3 Music Foundation Model(2 posts)→
More from Multimodal
- Creator breaks down psychedelic cosmic video techniques made with Seedance 2.5 — LudovicCreator · 2026-09-17
- NVIDIA's Axolotl3D does occlusion-aware 3D shape completion from multimodal inputs — NVIDIAAI · 2026-09-17
- fal teases Volume 2 of its generative media report — adamho · 2026-09-17
- 'Er Loop': a dark comedy AI short film retelling Plato's Myth of Er — Electrical_Toe8598 · 2026-09-17
- Runway's Fall 2026 drop aggregates Fish Audio, MiniMax, Cartesia, Flux models in one platform — tlakomy · 2026-09-17
- Runway launches Flux Video Edit to reshape objects and style in any video — runwayml · 2026-09-17