Qwen3.6-35B at Q2_XXS: A GPU-Poor Laptop Builds a Zelda-like RPG in 24 Minutes
ML-Future · reddit · 2026-09-03
On a Windows 11 laptop with just an i3 CPU, 8GB RAM and 0GB VRAM, the author ran Qwen3.6-35B-A3B-IQ2XXS GGUF locally via llama-server, sharing the full command (q80 KV cache, -ngl 0, --reasoning off). A single prompt for a Zelda-like 2D RPG produced a playable single-file HTML game in 24 minutes at 3 t/s. Being GPU-poor in 2026 is not so bad.
More from Infra
- Tencent Hunyuan 770B compressed from ~1.5TB to ~214GiB with mixed quantization — Aiden_Tech_Ai · 2026-09-03
- Cloudflare's Cache Transcoding shrinks cached text assets to ~1/3 with Zstandard — arpit_bhayani · 2026-09-03
- Microcenter shelf suddenly stocked with dozens of RTX 5090s — is the GPU shortage easing? — OvertaxedOne · 2026-09-03
- JapanFold offers 8 free open-source biology AI models on Japan-hosted inference infra — DavidBennett__ · 2026-09-03
- Squeezing Qwen 35B on an RX 6700 XT: a llama.cpp 100k-context tuning log — Loose_Doubt367 · 2026-09-03
- Open-sourced an experimental standalone DLSS 5 video player for neural rendering — 2600th · 2026-09-03