What Models Are People Running on 192GB Machines?

CentrifugalMalaise · reddit · 2026-07-19

A user with 192GB of RAM shares their experience running local LLMs.

They primarily run the Unsloth Dynamic UD-Q3KXL gguf for Qwen3.5-397B and mention using Claude to fix some llama.cpp issues related to hybrid recursive models, including:

They also list other LLMs they are interested in trying: GLM 4.7 357B, Deepseek V4 Flash 284B, Tencent Hy3 295B, Minimax M3 428B, Laguna M.1 225B, Minimax M2.7 229B, and Mimo 2.5 310B, while asking the community what others are running and how these compare to Qwen3.5-397B.

Original post →

More from Infra

Infra channel →