AI Safety Researchers Debate Limits of Monitoring Approach

AI safety researchers pushed back on Scott Alexander's claims, arguing that monitoring is far from solved and that chain-of-thought monitorability is increasingly failing.

2026-09-24 ~ 2026-09-24 · 2 related posts