Follow-up links the paper, code, and 240-second video comparisons
imjustnewatai · x · 2026-07-27
Follow-up with receipts for the 4-minute video model claim.
- Shares links to the 240-second video comparisons, the paper, and the code/model weights.
- Clarifies an important point: the model was trained using a 5-second window, but that does not mean the entire training dataset contained only 5 seconds of video.
- Useful as a source post, but it is still the same underlying event as the original announcement.
Related event: Video Model Trained in 5 Seconds Generates 4-Minute Clips(2 posts)→
More from Multimodal
- Claude wrote an entire song purely in code, no Suno involved — ctjlewis · 2026-09-23
- Opus 5.5 turns a single image into a Three.js game menu in one simple prompt — majidmanzarpour · 2026-09-23
- PixVerse's R2 world model goes hands-on: endless exploration, but compute limits cap play time — Xianbao_QIAN · 2026-09-23
- Local image generation tests: Qwen base model combined with a Flux refiner — freshstart2027 · 2026-09-23
- Reddit users find Qwen 2.1 a major disappointment for text-to-image — -becausereasons- · 2026-09-23
- Given a 3-hour budget and a one-line prompt, Opus 5.5 produced a full Kowloon horror short itself — rainbird · 2026-09-23