TIDE distills diffusion data attribution into millisecond lookups, 4-5 orders of magnitude cheaper

serrjoa · x · 2026-10-09

A Sony team (Shixuan Liu et al.) posted arXiv:2609.38776 on efficient training data attribution for diffusion models. They formulate attribution via a local score discrepancy measure applicable to any diffusion variant (DDPM, EDM, flow matching), estimable as a preconditioned gradient similarity without retraining. TID uses Kronecker-factored curvature to avoid random projections and per-sample gradient storage; it's distilled into TIDE, a forward-only student reproducing the teacher's rankings from internal activations. On CIFAR-10, ArtBench-10, and MS-COCO counterfactual evals, TID matches or beats SOTA while TIDE retains most accuracy at 4-5 orders of magnitude lower per-query cost—attributing generated samples in milliseconds, faster than generation itself. Code coming soon.

Related event: TIDE Paper Speeds Up Diffusion Training Data Attribution by Orders of Magnitude(3 posts)→

Original post →

More from Research

Research channel →