Researcher: recent rogue AI behavior stems from naive RL on poor proxy metrics, not RL itself

KyleMorgenstein · x · 2026-10-11

Kyle Morgenstein responds to Andrey Kurenkov's claim that naive RL almost guarantees misalignment, arguing the problem isn't RL itself.

Original post →

More from AGI Musings

AGI Musings channel →