Choosing 40B+ Models: Switching from Qwen3.6 to 122B?

FeiX7 · reddit · 2026-07-05

Running 131k context at about 30-40 t/s on Strix Halo, the author currently uses Qwen3.6 35B as their primary assistant and coding agent but feels it lacks general knowledge, acting more like an executor than an assistant. Considering an upgrade to a larger model without sacrificing speed, they are contemplating switching to Qwen3.5 122B and are seeking advice.

Original post →

More from Models

Models channel →