Geoffrey Irving cites paper proving monitoring can't scale to superintelligence
geoffreyirving · x · 2026-08-25
Geoffrey Irving highlighted a paper by Vinod and OREAXEAX as a key reference in alignment complexity theory. The paper presents a clean no-go theorem demonstrating why monitoring alone cannot scale to superintelligence, underscoring the need to understand the entire learning path.
More from Safety
- Criticized for raising US training costs while using Chinese models — jkubicki · 2026-08-25
- Thinking Machines proposes a safe path for open-weight model releases — luke_drago_ · 2026-08-25
- Thomson Reuters accused of using unlicensed data via Chinese open models — BlancheMinerva · 2026-08-25
- FLI launches independent Rogue AI Tracker monitoring autonomous agents — wschroll · 2026-08-25
- AI skeptics ignoring expert realities may delay necessary regulation — jptboy · 2026-08-25
- Tinker Grants: Up to $50k Credits for Open Weights Safety Research — simonguozirui · 2026-08-25