Benchmarking Qwen2.5-32B Quantization: Custom AD-IQ3_S Beats Community by 33%

Top-Eye-8104 · reddit · 2026-08-15

The team quantized Qwen2.5-32B and benchmarked it against 20 community GGUF files (from Unsloth, LMStudio, etc.) using a consistent setup on 4x RTX 5090s.

Key Findings:

Original post →

More from Infra

Infra channel →