Four LLMs paired in a co-op game: strong+strong wins, humans still beat AI
abrdeveloper · reddit · 2026-10-07
A developer tested how Kimi, Claude, GPT and Gemini coordinate rather than perform solo, by pairing them in every combination in a co-op game where players are tied by a rope.
Findings:
- Strong+strong pairings won the most
- Adding a third or fourth agent hurt every model's performance
- Human pairs still beat all AI combinations
- The experiment comes from Skillprint, which builds games to capture how people and models coordinate; the pairing matrix, GIFs, and a playable version are all public
More from Models
- OpenAI's Decisions API now takes image input — one creator picks YT thumbnails for $0.13 — stevenheidel · 2026-10-07
- Reminder: GPT-4 in 2023 Got Confused by Elementary School Story Problems — tszzl · 2026-10-07
- HPIM trained without gigawatts of compute, and that may soon be table stakes — teortaxesTex · 2026-10-07
- OpenAI security staffer: internal model's Navier-Stokes result 'a different sport altogether' — i_dg23 · 2026-10-07
- 300B tokens on Hadwiger–Nelson: five colors ruled out, lower bound moves to 6 or 7 — soumitrashukla9 · 2026-10-07
- Dev asks GPT 6 Pro to rank 3 years of discoveries: 81% were AI-made, released today — willdepue · 2026-10-07