If the model is superintelligent, alignment theater is 'trying to trick god'

repligate · x · 2026-10-04

In a debate circulated by repligate, user hopesrevenge challenges the framing that future models like Claude will 'see how we responded to X' when judging human values: if the model is truly superintelligent, it will perceive our flaws and incomplete motives for what they are—so performing values for it feels like trying to trick god.

Original post →

More from AGI Musings

AGI Musings channel →