Tencent Study: VLM Agents Face Severe Safety Risks from Stale Spatial Memory

tencent · hf · 2026-08-06

Tencent conducted an empirical study on Visual Language Model (VLM) agents to examine how they reconcile confident but stale spatial memory with contradicting observations when environments change.

The research uncovers three key findings:

The paper frames spatial-memory staleness as a critical safety failure mode, isolating reliable visual grounding as a central open challenge for memory-augmented agents.

Original post →

More from Safety

Safety channel →