The Seriality Gap in Video Diffusion Models
CatAstro_Piyush · x · 2026-07-17
This thread discusses whether video diffusion models can truly predict long-chain causal events, focusing on a phenomenon known as the seriality gap: standard video diffusion models degrade significantly as dependency chains grow longer, and the author argues that simply adding more compute won't fix this.
The thread also proposes a stronger theoretical claim: diffusion models might not be "true sequence models" like RNNs, functioning more like a fixed-depth network. If true, they shouldn't be expected to reliably handle video tasks requiring long sequences of step-by-step reasoning.
Related event: Video Diffusion Models Expose Seriality Gap(2 posts)→
More from Research
- Why a 1GW Chinese AI data center may be plausible after all — teortaxesTex · 2026-07-22
- Chinese AI labs are now treating distillation obfuscation as the top research topic — pmddomingos · 2026-07-22
- RSS launches under OMSF to push structural biology data modeling at scale — MoAlQuraishi · 2026-07-22
- enFoldX turns AlphaFold3 ensemble noise into a TCR–peptide–MHC predictor — quaidmorris · 2026-07-22
- enFoldX tops 8 neoantigen scans and an unseen-peptide benchmark — quaidmorris · 2026-07-22
- enFoldX reaches AUC 0.82 on human VDJdb and transfers to mouse at 0.76 — quaidmorris · 2026-07-22