Running Qwen3.8 UD-Q4_K_XL on M4 Pro; Q4 vs Q5 is only a ~3GB difference
TheZachMueller · x · 2026-08-18
A developer shares plans to run Qwen3.8's Unsloth Dynamic Q4KXL quant locally on an Apple M4 Pro, and asks the community whether anyone has compared Q4 vs Q5 — noting the two quantization levels differ by only about 3GB, worth weighing precision gains against memory usage.
More from Infra
- Merge Launches Model Router, Slashing Enterprise AI Spend by 75% — shensi · 2026-08-18
- Bank of America Targets Nvidia at $350, Bull Sees $1000 by 2028 — soumitrashukla9 · 2026-08-18
- tinygrad achieves 2.5 PFLOPS on AMD hardware using MXFP4 — AnushElangovan · 2026-08-18
- RDIMM ECC memory price surged $500 per 128GB in 2 weeks — HankYeomans · 2026-08-18
- T-Head RISC-V adds Day 0 support for Qwen-3.8, C950 hits 30 tps — pstAsiatech · 2026-08-18
- Nebius Expands Finnish AI Infrastructure to 455 MW — demian_ai · 2026-08-18