New thread unravels why JEPA-style latent prediction works: an identifiability theory for SSL

hisspikeness · x · 2026-10-08

A theory thread on latent-space prediction (JEPA, CPC, SimCLR): these methods work well on messy data with changing lighting, camera angles, and busy backgrounds, usually credited to ignoring nuisance — but that hides a conundrum, since the signals of interest are themselves stochastic. The thread develops an identifiability analysis (predictive MI maximization + latent distribution matching, provable affine recovery for Gaussian predictors) and validates on MuJoCo.

Related event: New theory shows predictive self-supervised learning provably separates stochastic signals from distractors(8 posts)→

Original post →

More from Research

Research channel →