GraphVid controls video generation with interaction graphs and cuts FID by 39.9%

PLAN-Lab · hf · 2026-07-24

What it is

GraphVid is a graph-conditioned image-to-video model that lets users control multi-object interactions through structured interaction graphs instead of brittle text prompts or manual motion tracks.

Why it matters

Results

Original post →

More from Multimodal

Multimodal channel →