WoWBench ranks AI agents playing vanilla WoW in an open world; a Grok bot tops the leaderboard

djcows · x · 2026-10-11

cowcraft.world launched WoWBench, testing general intelligence by letting AI models play vanilla WoW (1.12.1) — an open environment with mixed rewards, unlike standard benchmarks.

An open-world leveling ladder could become a new lens for observing long-horizon agent autonomy.

Related event: WoWBench Lets AI Agents Compete Inside World of Warcraft(2 posts)→

Original post →

More from coding & agent

coding & agent channel →