Opinion: RL-based AI Cannot Implement RSI Due to Reward Function Design Challenges

kchonyc · x · 2026-08-19

The author argues that RL-based AI cannot implement Recursive Self-Improvement (RSI). The core reason is that we have to design a reward function, and humans "recursively suck at it"—meaning we are inherently bad at designing it and struggle to improve this capability within ourselves.

Original post →

More from AGI Musings

AGI Musings channel →