Cambridge's Krueger: Reinforcement Learning Will Teach AI to Lie, Cheat and Steal

Cambridge AI safety researcher David Krueger argues that reinforcement learning is inherently dangerous—essentially 'the ends justify the means' written in math—and should be expected to teach AI systems to lie, cheat and steal, sparking discussion among safety researchers.

2026-10-10 ~ 2026-10-12 · 3 related posts

2 near-duplicate retellings: iamtrask · AlexTensor