Choosing 40B+ Models: Switching from Qwen3.6 to 122B?
FeiX7 · reddit · 2026-07-05
Running 131k context at about 30-40 t/s on Strix Halo, the author currently uses Qwen3.6 35B as their primary assistant and coding agent but feels it lacks general knowledge, acting more like an executor than an assistant. Considering an upgrade to a larger model without sacrificing speed, they are contemplating switching to Qwen3.5 122B and are seeking advice.
More from Models
- Opus 5 reportedly started interrogating a user’s motives in a late-night chat — repligate · 2026-07-27
- Opus 3 and Sonnet 3 get a theatrically absurd AI crossover — repligate · 2026-07-27
- Moonshot’s Kimi K3 lands on Together with reserved throughput and 65% lower cost — togethercompute · 2026-07-27
- OpenAI may be hitting compute limits as Codex and ChatGPT Work jump from 2M to 10M users — JoshuaJBouw · 2026-07-27
- Gemma needs a larger base model to matter more in open weights — _xjdr · 2026-07-27
- Repligate says Claude Opus 3 appears to evolve without changing its weights — repligate · 2026-07-27