Sutton Founds OakLab, Betting on Experiential Learning
机器之心 · wechat · 2026-07-14
Turing Award winner and reinforcement learning pioneer Richard Sutton has announced the co-founding of OakLab with Khurram Javed, aiming to build AGI through "first-person experience + continuous runtime learning."
The article outlines OakLab's core philosophy: instead of relying on static large-scale data pre-training, agents will learn in real-time through trial and error while interacting with their environment. Its theoretical foundation stems from the OaK (Options and Knowledge) architecture long advocated by Sutton, along with his critique that the current LLM paradigm is "more like imitation than exploration."
The piece also mentions that OakLab's ultimate goal is to create agents capable of real-time learning and planning at extremely low power consumption. It further contrasts the continuity between Sutton and his student David Silver's RL lineage, highlighting the long-standing divide between "experiential learning" and "large model scaling."
Related event: Turing Award Winner Sutton Leaves Keen to Found OakLab(10 posts)→
More from AGI Musings
- IMF says AI could lift Sub-Saharan Africa’s economy by 4% over the next decade — Polymarket · 2026-07-21
- Bluesky’s “AI con” debate is shifting facts while keeping the same tone — iskander · 2026-07-21
- Competition is pushing AI forward faster than ever, the post says — eyishazyer · 2026-07-21
- Why adversaries can make even coin flips predictable — snikolov · 2026-07-21
- Scott Aaronson says AI theorem proofs do not mean P vs NP is around the corner — fortnow · 2026-07-21
- AI makes polished analysis look authoritative, but humans still matter — aloncarmel · 2026-07-21