After Copilot went unlimited, one Reddit user is deciding whether to sell a two-GPU AI rig
tweetibird · reddit · 2026-07-27
Reddit user weighs keeping a local AI workstation after Copilot became unlimited at work
The poster built a local AI rig in September to run and learn models, including a setup with two RTX 6000 Pro Q-Max GPUs, 128GB of DDR5 RAM, a Ryzen 9950X, and two 4TB NVMe drives.
They used it for models like GPT-OSS 120B in Ollama, but after Microsoft made Copilot effectively unlimited for work and personal use, the machine has mostly sat idle. The user then bought a Mac mini M4 with 64GB RAM and has been running Qwen 3.6 35B A3B MLX 8-bit successfully for personal coding and homelab use.
The question is whether to sell the heavyweight rig or keep it for future local model needs once personal Copilot access becomes restricted again.
More from Infra
- Open-source profiler tracks every STT, LLM, and TTS call in self-hosted voice agents — mahimairaja · 2026-07-27
- SemiAnalysis says better memory and storage can beat a faster GPU in modern inference — rwang07 · 2026-07-27
- llama.cpp merges GLM-5.2-Vision support for local multimodal inference — QuixiAI · 2026-07-27
- Building an LLM server taught one author how hard self-hosting and reliability are — MaxChamp08 · 2026-07-27
- Cheap storage makes SCD Type 2 look obsolete, says a Meta-style data engineer — Zachly · 2026-07-27
- HuggingHack adds S3, MinIO, Ollama and vLLM dispatch in a self-hosted layer — TyedalWaves · 2026-07-27