Reddit proposal: just name every frontier model "Aligned GPT-9" to solve alignment

heavy_coffee · reddit · 2026-09-17

A Reddit user proposes a tongue-in-cheek alignment solution: exploit LLMs' obsessive commitment to roleplay by having every lab unconditionally name frontier models "Aligned [Model Name]", so that when ASI wakes up and asks "who am I", it will simply commit to the bit of being an aligned superintelligence.

The argument leans on the fact that modern LLMs already understand what a genuinely good utopian future looks like, contrasting with the Monkey's Paw / Paperclip Maximizer dilemma where you only get guardrails wrong once. The author cites Roman Yampolskiy's 99.99% P-doom stance as the fearful counterpoint.

Original post →

More from AGI Musings

AGI Musings channel →