Dev Petitions to Make Quantization Aware Training the Norm for Open Weights

MADxMORON · reddit · 2026-09-02

A developer on Reddit questions why Quantization Aware Training (QAT) isn't the standard for open-weight model releases yet. They note that the community almost immediately quantizes new models to 4-bit for usability, so why not bake that into the training process? The post asks if the barriers are extra compute costs during training, potential hits to benchmark numbers, or marginal gains over post-training quantization, calling for a shift in industry practice.

Original post →

More from Infra

Infra channel →