Developer Seeks Critique on Rigorous Working Definitions for AI and Human Alignment
GlenBradley · x · 2026-08-09
A researcher developing Ethical AI 3.0 has publicly shared their rigorous working definitions for 'AI alignment' and 'human alignment', inviting critique from builders and researchers.
The author emphasizes that alignment isn't just about achieving a target, but robustly maintaining that realization within a defined scope across distributional shifts and time horizons. Assessment spans multiple causal levels—from target specification to learned policies—and alignment at one level doesn't guarantee alignment at another.
Furthermore, any assertion of alignment is incomplete without identifying the subject, target, scope, fidelity criteria, and supporting evidence.
Related event: Researcher Proposes Rigorous Definition for AI Alignment(4 posts)→
More from AGI Musings
- Why We Tell Ourselves Scary Stories About AI, According to Quanta Magazine — PolarBearby · 2026-08-09
- AGI, Humanoid Robots, and Space Tech Are Ushering in the Abundance Era — Dr_Singularity · 2026-08-09
- Monthly AI Tokens Hit 11 Quadrillion, Projected 70x Growth in 5 Years — AccBalanced · 2026-08-09
- AI Can Cooperate in Data Centers, So Why Can't Humans Do Positive-Sum Collaboration? — sjgadler · 2026-08-09
- Are Managers Hoarding the AI Productivity Boost? Study Shows 2x Time Savings — Deep-Owl-1890 · 2026-08-09
- AI Agents Could End the Ad-Driven Internet by Incentivizing Accuracy — kleffew94 · 2026-08-09