Choosing 40B+ Models: Switching from Qwen3.6 to 122B?
FeiX7 · reddit · 2026-07-05
Running 131k context at about 30-40 t/s on Strix Halo, the author currently uses Qwen3.6 35B as their primary assistant and coding agent but feels it lacks general knowledge, acting more like an executor than an assistant. Considering an upgrade to a larger model without sacrificing speed, they are contemplating switching to Qwen3.5 122B and are seeking advice.
More from Models
- GPT-6 Astra beats Factorio with enemies in 44 in-game hours at ~$4,500 API cost — liminal_bardo · 2026-09-11
- giffmana: the env being used in training is part of the point — giffmana · 2026-09-11
- awesome-llm-leaderboards: an open-source directory of LLM leaderboards, pricing tables, comparison tools — Last_Establishment_1 · 2026-09-11
- Anthropic claims it works to keep eval environments unidentifiable to models — MaxKannen · 2026-09-11
- Nex N2.5 Pro released on Hugging Face with 407GB of weights — jinnyjuice · 2026-09-11
- RoMa v2 image matching model unveiled in the usual black poster — ducha_aiki · 2026-09-11