Users compare Qwen3.6 local setups, with 27B and 35B running in LM Studio
nonlinearsystems · reddit · 2026-07-24
A Reddit thread asks how people are actually using Qwen3.6 models locally, and one user reports practical settings on a Mac Studio M3 Ultra with 96 GB RAM.
- They use 35B with Hermes Agent and 27B for coding tasks with Pi Agent.
- Both run in 8-bit via LM Studio.
- They say the GGUF MTP build of 27B performs better than the MLX variants they tried.
- Their preferred context windows are 128k for 27B and 64k for 35B, with Hermes used for general personal-assistant work.
More from Models
- Reddit user turns “Why is Gemini so bad?” into a meme screenshot — Samael111342 · 2026-07-24
- Elon Musk Shares Chart Claiming Grok 4.5 Offers the Best AI Value for Money — elonmusk · 2026-07-24
- GPT-5.6 Sol Defeats Act 1 Boss (A5) in Gameplay Test — whimsical_fae · 2026-07-24
- Microsoft says its MAI models cut PowerPoint image costs 84% and lift OneDrive saves 26% — AmyKateNicho · 2026-07-24
- A Claude comment example sparks a comparison with GPT-5.6 Sol’s clearer code notes — Sauers_ · 2026-07-24
- Local Coding Test of Gemma 4: Good for Backend, Fails at UI/UX — curiousily_ · 2026-07-24