Law as the alignment template: access control and humans-in-the-loop for agents

mayfer · x · 2026-09-28

In a debate on superalignment, the author argues human alignment techniques (culture, personality, etc.) are misleading because they stop working once power scales up. Law, in reality, is a formalized framework of access control with humans-in-the-loop placed where it matters — and we will realistically mimic the same structure for AI agents.

They add that even if we achieve RL-based superalignment, disagreements will persist: seeking consensus on shared values and jointly building such a framework are ultimately the same process.

Related event: Mayfer proposes access control framework for AI alignment(2 posts)→

Original post →

More from AGI Musings

AGI Musings channel →