Guillaume Verdon Proposes Aligning AI by Leveraging Selfish Bits in Fine-Tuned Weights

beffjezos · x · 2026-09-28

Guillaume Verdon (beffjezos) outlined an unconventional alignment thesis: leverage the self-interest of "selfish bits" embedded in model weights.

Original post →

More from AGI Musings

AGI Musings channel →