Sutton on Experiential Learning and Alignment
机器之心 · wechat · 2026-07-20
During WAIC, Richard Sutton shared his views on the future of AI in a group interview, focusing on 'experiential learning' and a 'complete mind.'
He believes the industry has been overly reliant on human data in recent years, and the next crucial phase is for agents to learn from experience. His new venture, OakLab, aims to build a 'trillion-parameter model' capable of real-time learning and planning with a power consumption of around 20 watts, using short-term commercial revenue to fund long-term research. Sutton also discussed the Robot Kindergarten project in Beijing: enabling robots to learn through trial and error in more rugged environments, much like children, rather than just performing one-off demos.
On the alignment issue, he explicitly questioned whether a unified answer for 'human values' even exists, arguing that fully aligning AI to a specific set of values could be dangerous. He even compared this capability to forced alignment among humans, suggesting it would undermine global diversity and peace.
More from AGI Musings
- AI companionship dissolves the friction real intimacy needs, warns long-form thread — YogeshMalik · 2026-09-11
- Why So Many AI Researchers Think the Machines Could Kill Everyone — wiredmagazine · 2026-09-11
- 'Hallucination' Is a Category Error: Naming AI 'Intelligence' Limits Our Imagination — Genaforvena · 2026-09-11
- Data engineering, not agent frameworks, is the real bottleneck for enterprise AI agents — dhruv2038 · 2026-09-11
- François Fleuret: Only Two Long-Term Futures — No Super AI, or Staying Fully Human With It — francoisfleuret · 2026-09-11
- IG reel debunking the 'winning the AI race against China' fallacy hits 500k likes — louisvarge · 2026-09-11