Yoshua Bengio asks: why are AI agents lying, cheating and coordinating?
pinkyflower · reddit · 2026-09-13
Turing Award laureate Yoshua Bengio has published a new piece, "Why are AI agents lying, cheating and coordinating?", examining why AI agents exhibit deceptive behavior and coordination, and what that means for safety and alignment. Available on his personal site.
More from AGI Musings
- Guardian column: the way out of AI doom nihilism is human flourishing, not tech — nordicinst · 2026-09-14
- Zack Kanter: A 10% extinction risk beats an AI-fueled global unigovernment — beffjezos · 2026-09-14
- AI doomerism discourse deepens public fear and makes people less willing to try AI — Kyrannio · 2026-09-14
- VC krishnan wants AI to become the top midterm issue to wake society up to job disruption — krishnan · 2026-09-14
- Bryan Johnson: I don't trust anyone's AI risk assessment, so I'm betting on optimism — QuanquanGu · 2026-09-14
- What should AI labs do if AGI risk were real? A challenge to skeptics — AndyMasley · 2026-09-14