Richard Sutton praises PhD thesis on robots that keep learning after deployment

RichardSSutton · x · 2026-10-04

Richard Sutton highlights Gautham Vasan's new PhD thesis "Robots That Learn on the Fly Through Real-World Interaction," which he says is exceptionally well done. Chapter 6 introduces AVG, a new streaming RL actor-critic algorithm distinct from the Stream-X family.

The thesis attacks the prevailing train-then-freeze paradigm in robot learning: policies are fixed after deployment, so degradations require engineers to collect data and retrain. It instead enables robots to learn from their own runtime experience, improving over their operational lifetime and adapting to unforeseen conditions.

Original post →

More from Embodied

Embodied channel →