Self-play RL finally cracks Brood War — with a 300M-parameter model

syhw · x · 2026-09-25

Google DeepMind researcher syhw says his 2016 quest to solve Brood War from scratch with self-play RL has been completed — by a 300M-parameter model, he announced with a link. A striking contrast to AlphaStar's nine-figure training budgets, spotlighting efficient RL training.

Original post →

More from Fun

Fun channel →