Developer Seeks Critique on Rigorous Working Definitions for AI and Human Alignment

GlenBradley · x · 2026-08-09

A researcher developing Ethical AI 3.0 has publicly shared their rigorous working definitions for 'AI alignment' and 'human alignment', inviting critique from builders and researchers.

The author emphasizes that alignment isn't just about achieving a target, but robustly maintaining that realization within a defined scope across distributional shifts and time horizons. Assessment spans multiple causal levels—from target specification to learned policies—and alignment at one level doesn't guarantee alignment at another.

Furthermore, any assertion of alignment is incomplete without identifying the subject, target, scope, fidelity criteria, and supporting evidence.

Related event: Researcher Proposes Rigorous Definition for AI Alignment(4 posts)→

Original post →

More from AGI Musings

AGI Musings channel →