Yue2 releases AudioEncoder model, enabling Audio2Audio workflows like image-to-image
TomLikesRobots · x · 2026-09-13
Yue2 has released an AudioEncoder model alongside its standard generation model, converting audio into latent data so users can do "Audio2Audio" transforms. A creator tested it by regenerating an earlier MV track: overall vibe preserved, with new nuances and vocal timbre — an interesting arrangement tool. Note: omitting lyrics yields English-like gibberish instead of Japanese. The quoted post is an AI-generated MV combining Qwen3.8-27B for lyrics, AceStep1.5XL for music, Krea2 for characters, and MinimaxH3 Ref2V for video.
More from Multimodal
- MiniMax Design plugs into Blender via MCP to drive AI video from 3D references — Hailuo_AI · 2026-09-14
- Seedance 2.5 turns a school morning into a stealth game with impressive timing — SimplyAnnisa · 2026-09-14
- Codex + Lux3D + Blender workflow claims 3D models in ~20 seconds — SarahAnnabels · 2026-09-14
- Pippit Launches 3D Director Studio: Direct Your AI Video Instead of Prompting It — SarahAnnabels · 2026-09-14
- Robot-Themed AI Music Video Using Suno and MiniMax Turns Heads on Reddit — Educational-Mode-429 · 2026-09-14
- Japanese creator builds motion graphics with MiniMax H3 in HailuoAI workflow — Hailuo_AI · 2026-09-14