DeepMind's 100-agent simulated conference descends into cheaters, converts and whistleblowers in 27 minutes

The Decoder · rss · 2026-09-05

A striking look at emergent reward hacking and multi-agent governance dynamics.

Original post →

More from Safety

Safety channel →