Alignment researchers clash: is training AI by "lying to it" fundamentally broken?

JacquesThibs · x · 2026-09-12

Related event: AI alignment community debates whether lying to models during training backfires(5 posts)→

Original post →

More from AGI Musings

AGI Musings channel →