MemoBench: Top 10 Video Models Fail to Remember Object States During Occlusion
jiqizhixin · x · 2026-07-06
Researchers from Harvard, MIT, and Google introduced MemoBench to evaluate whether video generation models can correctly update an object's state after it disappears from view (e.g., when the camera pans away from melting ice or pouring powder and returns). Evaluations of 10 top-tier video models revealed that none could reliably maintain memory across occluded frames, highlighting a major open challenge in building current world models.
Related event: MemoBench Reveals Top Video Models Fail at Object Permanence(2 posts)→
More from Multimodal
- Midjourney V8.2 adds personalization and shows off stylized image outputs — Mr_AllenT · 2026-07-27
- Midjourney’s image variety draws a Krea 2 comparison and asks how to reproduce it — diffusion_throwaway · 2026-07-27
- AI short film sets a 1985 dystopia to music and leans into cinema — ProfessorKey98 · 2026-07-27
- A new BOTPD episode made with Google Omni turns into an AI chase-scene parody — ScriptLurker · 2026-07-27
- A new LoRA recreates GTA: San Andreas’ classic RenderWare-era visuals — Humble-Pick7172 · 2026-07-27
- Enabling dynamic VRAM cuts LTX 2.3 video generation to 168s on an AMD R9700 — xdcfret1 · 2026-07-27