MiniMax Releases Open-Weights Music Model Music3 for 5-Minute Full Songs

BanghuaZ · x · 2026-08-14

MiniMax has introduced MiniMax-Music3, a new open-weights music generation model. Conditioned on lyrics and detailed music descriptions, it can generate complete songs up to five minutes long with structural coherence, expressive vocals, and stable long-form audio quality.

The architecture combines several components:

The model outputs 32 kHz, 16-bit stereo WAV audio. It also supports fine-grained control, allowing users to use structured tags (e.g., verse, chorus, bridge) and metadata (genre, BPM, emotion) for precise customization. The model is now available on Hugging Face and can be served using SGLang.

Related event: MiniMax Open-Sources Music 3 Text-to-Music Model(14 posts)→

Original post →

More from Models

Models channel →