Open-source YuE2 beats Suno v5/v6 on WildSongBench with symbolic score planning
solyarisoftware · x · 2026-09-15
A repost argues open-source music models have hit an inflection point, with YuE 2.0 matching or beating closed alternatives like Suno, ElevenLabs and Lyria 3.5 in certain capabilities.
Key points:
- Approach: YuE2 unifies symbolic and audio generation — it plans an editable ABC-notation score first, then renders vocals and accompaniment, supporting creation, covers, and agentic conversational editing.
- Specs: 3.59B parameters, 28 layers, from M·A·P, Tokenwave.AI, MBZUAI, and ACE STUDIO.
- Results: best-of-8 scores 6.9632 on WildSongBench (192 prompts), the highest of 17 evaluated settings; Suno v5 scores 6.8721, Suno v6 and v6 Wild score 6.5562 and 6.4195.
- Take: remaining gaps in controllability and coverage are marginal and fixable via local fine-tuning, while closed-source music models show diminishing returns.
Related event: Open-Source Music Model YuE2 Beats Suno v5/v6 on WildSongBench(4 posts)→
More from Multimodal
- Cartesia explains why benchmarking TTS is extremely hard: no single number captures voice quality — saranormous · 2026-09-16
- YuE2 music sampling only hits ~7 tokens/s on RX 9070 despite mostly idle VRAM — Prestigious-Kick7291 · 2026-09-16
- Pixio launches AE plugin: chat agent builds native layers and keyframes in your timeline — tsi_org · 2026-09-16
- Midjourney style ref + Gemini photorealism: a two-step image workflow — michaelrabone · 2026-09-16
- Seedance 2.5 recreates a GTA-style stealth mission with flawless phone tracking — SimplyAnnisa · 2026-09-16
- Nex-N2.5 Pro builds an interactive 3D mechanical garden by iterating on its own output — nikola_mr64990 · 2026-09-16