Follow-up link on humans and VLMs producing similar memory-based reconstructions

ducha_aiki · x · 2026-09-09

A follow-up link post by duchaaiki on the same thread: humans and VLMs produce surprisingly similar outputs when reconstructing an image from memory or text description, tied to Jakob Engel's remarks on world models.

Related event: DeepMind's Jakob Engel on World Models: Egocentric Multimodal Data Is the Future(4 posts)→

Original post →

More from Research

Research channel →