JEPA-like world models collapse on distractors and natural video, researchers report

inductionheads · x · 2026-09-27

@bonniesjli responded to a question about whether JEPA has ever been validated on standard RL tasks: the team compared against a Momentum Prediction baseline (similar to JEPA) and found it comparable in clean settings but collapsed under distractor and natural video conditions. The takeaway: JEPA-like world models need more work to function well on these tasks. The original questioner asked whether JEPA has ever been shown to work on DM Control or benchmarked head-to-head against standard model-based RL — this reply suggests such evidence remains scarce.

Original post →

More from Research

Research channel →