Multi-Agent LLMs Struggle to Explore Each Other

omarsar0 · x · 2026-07-19

This paper addresses the issue of how LLM multi-agents can effectively "explore each other." The authors point out that the default assumption—placing sufficiently capable models together will naturally result in good collaboration—does not hold true. Modern LLM agents often fall into myopic, polarized interaction patterns, leading to poor coordination and higher regret values.

The paper formalizes this phenomenon as the Multi-Agent Exploration problem: in partially observable stochastic games, agents must infer each other's capabilities through interaction and identify more effective collaboration strategies. To address this, the authors propose MACE, a lightweight framework that explicitly promotes exploration through structured peer selection. Results show that MACE significantly improves exploration behavior under both context diversity and parameter diversity settings, leading to better downstream task performance.

Original post →

More from coding & agent

coding & agent channel →