Qwen3.6-35B-A3B runs well on an RTX 4070, dev reports with UD-Q4_K_XL quant
haydendevs · x · 2026-09-22
Developer haydendevs shares first-hand experience that Qwen3.6-35B (MoE with 3B active parameters) runs pretty well on his RTX 4070, using the UD-Q4KXL quantization with MTP — a useful feasibility data point for running the new open-weights model on consumer GPUs.
Related event: Qwen3.6-35B Runs Smoothly on a Single RTX 4070(3 posts)→
More from Infra
- Banks reportedly halting compute lending as AI credit crunch begins — citrini · 2026-09-22
- California signs seven bills making AI data centers pay their own utility costs — The Verge AI · 2026-09-22
- O'Reilly builds a working data vocabulary for the semantic era, from warehouses to ontologies — rseroter · 2026-09-22
- UK's most powerful government AI supercomputer cost £225m — same as one bridge — emax · 2026-09-22
- exe.dev wins over developers: SSH into root VMs, plus the underrated Shelley coding agent — davidcrawshaw · 2026-09-22
- NVIDIA hosts trilateral meeting as Artificial Analysis becomes Korea sovereign AI evaluator — ArtificialAnlys · 2026-09-22