GenCine was trained on academic compute: full technical details now on the project page

anand_bhattad · x · 2026-10-03

Closing the GenCine thread: all training ran on academic compute (JHU DSAI support). The project page details the pipeline—depth-estimated colored point clouds, Blender-authored motion handles, per-frame background/foreground XYZ and identity maps encoded by a frozen Wan VAE, conditioned into the frozen diffusion model via a half-resolution side branch and LoRA adapters, with only those components trained.

Related event: GenCine: Single-Image Video Generation with 3D Motion Control(3 posts)→

Original post →

More from Multimodal

Multimodal channel →