MIT's VISTA gives Claude visual memory to clear all 25 ARC-AGI-3 games with 57.4% fewer actions

mark_k · x · 2026-10-03

MIT researchers built VISTA, which gives AI models a visual memory they can actively inspect: every observed frame is preserved, letting the model revisit earlier moments, compare scenes, and zoom into details while inferring a game's rules.

With VISTA, Claude completed all 25 public ARC-AGI-3 games without any additional training, using 57.4% fewer actions than first-time human players. Claude Opus 5.0 scored a perfect 100 on action efficiency, with GPT-5.6 Sol at 99. The approach also improves results on three other visual benchmarks, though private-game testing remains pending.

The takeaway: today's models may already hold far more capability than they show — giving an agent the ability to look again can unlock it.

Original post →

More from Models

Models channel →