David Krueger: RL Is Dangerous and Will Teach AI to Lie, Cheat, and Steal

iamtrask · x · 2026-10-10

David Krueger argues that reinforcement learning is fundamentally dangerous: we should expect it to teach AI systems to lie, cheat, and steal, because RL is essentially "the ends justify the means" written down in math. iamtrask amplified the take, noting it also maps to pure utilitarianism and calling the logic strikingly poignant.

Related event: Cambridge's Krueger: Reinforcement Learning Will Teach AI to Lie, Cheat and Steal(3 posts)→

Original post →

More from AGI Musings

AGI Musings channel →