Point-by-point rebuttal of Scott Alexander: AI monitoring far from solved

austinc3301 · x · 2026-09-24

In the debate sparked by Scott Alexander's 'Hell' post, austinc3301 lists concrete objections: monitoring is far from solved (collusion, multi-agent misalignment, opacity); CoT monitorability looks weaker by the day and companies are incentivized to compromise it; opportunism remains a risk under monitoring; interpretability isn't intractable; and lower per-agent intelligence needs for agent swarms won't stop companies from maxing out individual agent intelligence — they're already trying.

Original post →

More from AGI Musings

AGI Musings channel →