AI Safety Researchers Debate Limits of Monitoring Approach
AI safety researchers pushed back on Scott Alexander's claims, arguing that monitoring is far from solved and that chain-of-thought monitorability is increasingly failing.
2026-09-24 ~ 2026-09-24 · 2 related posts
- Point-by-point rebuttal of Scott Alexander: AI monitoring far from solved — austinc3301 · 2026-09-24
- AI safety researcher pushes back: CoT monitorability failing, monitoring far from solved — austinc3301 · 2026-09-24