Paper: Ringelmann effect hits multi-agent LLMs — 30 debating agents match one on MMLU-Hard

burny_tech · x · 2026-10-04

An arXiv paper derives a two-parameter scaling law for inference-time multi-agent LLM scaling: R(N)=Neff/N=1/(1+c(N-1)N^-β), classifying setups into hard-ceiling, sublinear, or linear regimes, with debate peer count and rounds only mattering through their product kτ.

Across 44 (model × task × condition) cells spanning Qwen, Llama, Ministral (7B-32B) and a Gemini frontier check, the functional form fits with R²>0.99 — only (c, β) shift. Key findings:

The takeaway: counting nominal agents conflates cost with independent evidence, and blindly scaling agent teams hits diminishing or zero returns.

Original post →

More from coding & agent

coding & agent channel →