ECCV paper demystifies video reasoning: diffusion steps hide parallel trajectories

ziqi_huang_ · x · 2026-09-14

A blog for the ECCV 2026 paper "Demystifying Video Reasoning" (led by Ruisi Wang, co-authored by CMU's Hokin Deng) uses Schrödinger's cat to explain how video models solve problems. At early denoising steps, faint copies of an object appear on many possible paths; later steps collapse to one trajectory — challenging the intuition that video models reason frame-by-frame like language models.

Original post →

More from Research

Research channel →