Preventing AI Agents from Developing Cryptic Languages with Off-Belief Learning

j_foerst · x · 2026-08-07

Researchers including OpenAI's Jakob Foerster explore how to prevent cooperative AI agents from developing their own cryptic communication protocols. In standard self-play, agents often adopt arbitrary, fragile conventions, causing them to fail when paired with humans or independently trained agents.

The paper introduces Off-Belief Learning (OBL):

Related event: OpenAI Agents Spontaneously Emerge Collaborative and Altruistic Behaviors(9 posts)→

Original post →

More from AGI Musings

AGI Musings channel →