'Aligned is as aligned does' only works retrospectively; we need an intensional definition

zetalyrae · x · 2026-09-11

The author argues that many people hold an "extensional" definition of alignment—"aligned is as aligned does"—but this can only be applied retrospectively. It doesn't help prospectively answer: is this AI aligned? That requires an intensional definition, which is much harder, since you must first clarify a stack of concepts: what is an "agent," "intelligence," "goals," and what separates agent from environment, principal from agent.

Related event: The alignment definition problem: extensional views are retrospective only(3 posts)→

Original post →

More from AGI Musings

AGI Musings channel →