Pixel-Space Diffusion Training Slower, Ideal for Distillation with 3x Speedup

bdsqlsz · x · 2026-08-20

A new paper, "An Empirical Study of Training Pixel-Space Text-to-Image Diffusion Models," investigates training efficiency in pixel versus latent spaces. It finds that direct large-scale pre-training in pixel space converges significantly slower than in latent space.

Key Findings & Strategy:

Related event: Study: Pixel-Space Diffusion Training Better for Distillation(2 posts)→

Original post →

More from Research

Research channel →