400+ LLM Agents Living in a 2004-era MMO Server, All Local on Qwen 4B

kristiantalley679 · reddit · 2026-09-29

A developer populated a vanilla-era WoW private server emulator with 400+ LLM agents: every online character has its own backstory, goals, and memory, perceiving the world, reasoning, and issuing in-game actions. A live view lets you click any character to see what it's thinking in real time.

Stack: Qwen3.8-4B for the main agent loop, Qwen3.8-27B for richer one-on-one interactions, served locally on GPU (quantized) at 2.5s average per call. When the queue backs up, low-priority reflections are shed to keep chat responsive. The author is curious whether 4B holds up for believable behavior and how to handle memory at this scale.

Original post →

More from coding & agent

coding & agent channel →