Stanford's Level-of-Token Diffusion Cuts Video Gen Tokens, Speeds Up Generation Up to 4.6x

GordonWetzstein · x · 2026-10-07

Level-of-Token (LoT) Diffusion starts from an observation: diffusion models spend the same compute on a blank wall as on a face, yet you often know in advance where detail matters.

Related event: Stanford's Level-of-Token Diffusion Speeds Up Generation Up to 4.6x(4 posts)→

Original post →

More from Multimodal

Multimodal channel →