Hide Model Names From Agents: Labels Cost 55% More Tokens and Drop Success to 81%

alex_verem · x · 2026-10-08

Detailed writeup of the Sapienza University study: in group games with 9-25 agents, exposing model names caused factions regardless of actual model capability — fake names and color labels produced the same clustering. A Qwen agent abandoned a correct answer for a same-family agent's wrong one, citing "Qwen models are usually reliable." Labeled groups needed 30% more rounds and 55% more tokens, with success dropping from 96% to 81%. Recommendation: orchestrators may keep model info for routing/debugging but should never expose it to the agents.

Related event: AI Agents Spontaneously Form In-Groups by Model Family, Study Finds(2 posts)→

Original post →

More from Research

Research channel →