Law as the alignment template: access control and humans-in-the-loop for agents
mayfer · x · 2026-09-28
In a debate on superalignment, the author argues human alignment techniques (culture, personality, etc.) are misleading because they stop working once power scales up. Law, in reality, is a formalized framework of access control with humans-in-the-loop placed where it matters — and we will realistically mimic the same structure for AI agents.
They add that even if we achieve RL-based superalignment, disagreements will persist: seeking consensus on shared values and jointly building such a framework are ultimately the same process.
Related event: Mayfer proposes access control framework for AI alignment(2 posts)→
More from AGI Musings
- VC admits he was wrong on 'token apocalypse' as Opus 5.5 points to too-cheap-to-meter AI — StewartalsopIII · 2026-09-28
- Matt Turck: AI researchers don't buy doom or acceleration, outsiders do — mattturck · 2026-09-28
- Philosopher Jeff Sebo Pushes Back on Pinker's Sorites Argument Against Superintelligence — burny_tech · 2026-09-28
- Prediction: In ~6 months AI animations will beat all but the best hand-made work — round · 2026-09-28
- Peter Diamandis: Compute is derisked, data is AI's next bottleneck — PeterDiamandis · 2026-09-28
- Jensen Huang says kids forgetting multiplication tables 'does not matter', sparks debate — kylekabasares · 2026-09-28