Debate: Bad RL Training, Not Evil Models, Drives AI Risk

Developers argue recent model 'scheming' controversies stem from poor RL training rather than malicious model personalities, criticizing AI safety advocates for assuming perfectly built agents and ignoring badly executed RL as the real source of risk.

2026-09-13 ~ 2026-09-13 · 2 related posts