Unsloth requested to re-quantize older Qwen models using UD 3.0
Fancy-Snow7 · reddit · 2026-08-27
A Reddit user requested the Unsloth team to re-quantize Qwen 72B and 27B models using the UD 3.0 standard. The user noted that UD 3.0 is a massive improvement over UD 2.0, claiming that Q3 quantization with UD 3.0 yields results similar to Q4 with UD 2.0, bringing efficiency gains to older model weights.
More from Infra
- GPU Cloud Showdown: RunPod vs. Vast vs. Nebius and the Env Fragmentation Tax — big-in-jap · 2026-08-27
- Data centers are 'intelligence factories': Energy is the fuel of the new industrial age — NinaDSchick · 2026-08-27
- Integrating Warp Factory run costs into Slack for model routing optimization — vikvang1 · 2026-08-27
- Cloudflare Workers testing granular permissions for agents — dinasaur_404 · 2026-08-27
- Weaviate 1.39 ships Boost API and MMR to GA, adds 4-bit RQ quantization — CShorten30 · 2026-08-27
- Phala Confidential AI bills 61.5B tokens in 24h, DeepSeek leads at 54% — bgmshana · 2026-08-27