ProgramBench multi-agent eval: Opus 5.5 fastest with a 5-agent team, Sonnet 5.5 with subagents

jyangballin · x · 2026-09-29

In a multi-agent evaluation on ProgramBench, the fastest configurations differ by model: Sonnet 5.5 performed best with subagents, while Opus 5.5 hit top speed using a 5-agent team. A useful data point that optimal agent orchestration is model-specific.

Original post →

More from coding & agent

coding & agent channel →