Researchers Warn: Opaque Multi-Agent RL at Labs Creates an Unstudiable Gap

dhadfieldmenell · x · 2026-09-27

Blanche Minerva voices concern that the research community doesn't know what training methods frontier labs are using, making the practices impossible to study. Tom Hadfield-Menell agrees and adds that multi-agent RL is almost certainly incredibly GPU-hungry, widening the gap between academic researchers and frontier labs in both information and compute.

Original post →

More from AGI Musings

AGI Musings channel →