An alignment framework arguing "English constitution" safeguards fail — start with an ethical topology of training data

GlenBradley · x · 2026-09-08

GlenBradley responds to a debate on alignment failure modes, arguing that bolting an English-language constitution onto cognition is insufficient: a capable intelligence can change representations (English → conlang → latent ontology → successor architecture), so any safety property that vanishes under a representation change was never real.

His layered proposal:

He sees "high-protein data" as one intervention in a larger problem — better source material changes the prior.

Original post →

More from AGI Musings

AGI Musings channel →