A new alignment glossary draft asks whether human and neutral definitions should split

GlenBradley · x · 2026-08-04

A user publishing an “Ethical AI 3.0” draft asks for hard feedback on how to define alignment, misalignment, and the alignment problem.

They propose four working definitions:

The post specifically asks whether the human-neutral split is useful, what dimensions are missing, and how to phrase the definitions so they survive adversarial and long-horizon edge cases.

Original post →

More from AGI Musings

AGI Musings channel →