Optimized Dual 3090 Quantization of Qwen3.8-27B Released

luedtek · reddit · 2026-08-15

A community release of an optimized Qwen3.8-27B quantization (INT8 W8A16 MTP) specifically tuned for dual RTX 3090 setups is now available on Hugging Face.

Original post →

More from Infra

Infra channel →