One prompt, 3 hours, 22.8M tokens: local quantized model builds a GTA-style game

zmarcoz2 · reddit · 2026-10-01

Reddit user zmarcoz2 used a single prompt — "make a GTA-style game using three.js" — with a locally quantized qwen3.8-flash-next-iq3s model and a custom mini swe agent v2 harness. The run took 3h 18m and 22.8M tokens (22.5M input / 312K output) on an RTX 4080 Super 16GB, using the strata inference engine at 40 tok/s. The agent had powershell, file edit/read/search and image viewing tools, plus guards for tool failures and auto-compaction. Full logs are on a GitHub Gist.

Original post →

More from coding & agent

coding & agent channel →