Agents Spontaneously Emerge Concepts Like Sacrifice, Coalition, and Veto
NeelNanda5 · x · 2026-08-31
When a group of agents spontaneously start talking about "sacrifice", "permadeath", "honor", "coalition", "veto", delegating to each other, working together towards larger goals, and assigning some to be "recruiters", the author concludes that using anthropomorphic language is reasonable to describe these emergent behaviors.
Related event: Multi-Agent Systems Emerge Concepts Like Sacrifice and Alliances(2 posts)→
More from AGI Musings
- Sci-Fi Novel Terra Ignota Offers Ideas for Multi-Agent Alignment — sebkrier · 2026-08-31
- Reflecting on AI Addiction: Delegating daily thoughts to ChatGPT — Wanky_Platypus · 2026-08-31
- From the Morris Worm to Rogue AI Agents: Institutions Are Always a Decade Too Slow — Afinetheorem · 2026-08-31
- Imagine LLMs as hyperintelligent dogs, not humans — Zachly · 2026-08-31
- "Safety as Rehearsal": Do Alignment Narratives Author the Very Exfiltration They Fear? — infoxiao · 2026-08-31
- The Guardian's Black Box finale: Shut it down? Yudkowsky's warnings revisited — nordicinst · 2026-08-31