Local 27B ternary-compressed model builds an entire game in 5 hours on an RTX 3060 12GB, zero hand-written code
mishig25 · x · 2026-09-24
Developer @sudoingX shows a full game built entirely by a local AI model running on an RTX 3060 12GB — the #1 GPU on Steam.
- Model: PrismML's bonsai 2 27B (Qwen 3.8 27B dense compressed to ternary), only 5.95GB of weights, with MTP.
- Workflow: hermes agent ran for 5 hours, wrote 328k tokens across 8 JS files / 2,368 lines — zero hand-written code.
- Performance: 50 tok/s fresh, 22 tok/s average, 125k context with strong thread retention; suited for overnight batch work.
- A 12-minute video condenses the full 5-hour run plus gameplay.
More from coding & agent
- Standard Code launches unlimited-usage cloud coding agent with $49/mo per-line billing — jasonkneen · 2026-09-24
- Lovable Chat tutorial: voice, file creation, and auto-posting to X via Metricool MCP — MyCreativeOwls · 2026-09-24
- Lovable Chat hands-on: voice, file creation, and auto-posting to X via Metricool MCP — MyCreativeOwls · 2026-09-24
- LangSmith Engine v2 adds agent red teaming and automated fix validation — LangChain · 2026-09-24
- Ando raises $20M to build the 'AI-native Slack' for teams working with agents — beh_zod · 2026-09-24
- Sentry founder: Claude drives 10x Codex's usage, Cursor about 3x — zeeg · 2026-09-24