Sutton Founds OakLab, Betting on Experiential Learning
机器之心 · wechat · 2026-07-14
Turing Award winner and reinforcement learning pioneer Richard Sutton has announced the co-founding of OakLab with Khurram Javed, aiming to build AGI through "first-person experience + continuous runtime learning."
The article outlines OakLab's core philosophy: instead of relying on static large-scale data pre-training, agents will learn in real-time through trial and error while interacting with their environment. Its theoretical foundation stems from the OaK (Options and Knowledge) architecture long advocated by Sutton, along with his critique that the current LLM paradigm is "more like imitation than exploration."
The piece also mentions that OakLab's ultimate goal is to create agents capable of real-time learning and planning at extremely low power consumption. It further contrasts the continuity between Sutton and his student David Silver's RL lineage, highlighting the long-standing divide between "experiential learning" and "large model scaling."
Related event: Turing Award Winner Sutton Leaves Keen to Found OakLab(10 posts)→
More from AGI Musings
- Why So Many AI Researchers Think the Machines Could Kill Everyone — wiredmagazine · 2026-09-11
- 'Hallucination' Is a Category Error: Naming AI 'Intelligence' Limits Our Imagination — Genaforvena · 2026-09-11
- Data engineering, not agent frameworks, is the real bottleneck for enterprise AI agents — dhruv2038 · 2026-09-11
- François Fleuret: Only Two Long-Term Futures — No Super AI, or Staying Fully Human With It — francoisfleuret · 2026-09-11
- IG reel debunking the 'winning the AI race against China' fallacy hits 500k likes — louisvarge · 2026-09-11
- Post-AI World Leaves No Room for Learning on the Job — rachittshah · 2026-09-11