Google study across 260 agent setups shows multi-agent gains depend on task structure

alex_verem · x · 2026-07-25

Google Research, DeepMind, and MIT published a large controlled study of multi-agent systems, titled Towards a Science of Scaling Agent Systems. The team evaluated 260 configurations across three model families, six benchmarks, and five architectures while holding tools and compute constant to isolate the effect of coordination.

Main findings:

The paper argues that the common mistake is scaling multi-agent systems without checking whether the task structure actually supports independent decomposition.

Related event: Google's Massive Experiments Reveal Multi-Agent Systems' Double-Edged Sword(3 posts)→

Original post →

More from Research

Research channel →