Two AI Models Autonomously Generate a Complete Music Video
petrusenko_max · x · 2026-07-17
A post demonstrates a real test of automatically generating a complete music video for "Uptown Funk":
- GPT-5.6 Sol uses an image-to-video pipeline with a budget cap of $25.
- Claude Fable 5 uses a text-to-video approach and successfully generates outputs across two different budget tiers.
- Results show that higher budgets yield more shot footage, though models vary in resolution and spending strategies.
Related event: AI Models Tested on Generating Full Music Videos Under $100(2 posts)→
More from Multimodal
- Pablo Stanley shares a full AI video workflow using ChatGPT, Gemini, Runway and CapCut — jdjohnson · 2026-07-21
- Meta AI text input now lets users interleave images with text — ezyang · 2026-07-21
- ShotPlan adds learnable planning tokens for cinematic multi-shot video generation — Tele-AI · 2026-07-21
- Same prompt, Seedance 2 and Grok are compared on cinematic transformation output — LudovicCreator · 2026-07-21
- CG Chefs Showcases Retro Anime Style AI Video Generation — nicolascraske · 2026-07-21
- Night-party video demo uses Seedance 2.0, timecode prompts and 4K upscaling — gen_ericai · 2026-07-21