DeepSeek 2.52-bit quantization holds up surprisingly well in testing

nomorebuttsplz · reddit · 2026-08-24

User testing found that a heavily quantized 2.52-bit EXL3 version of DeepSeek Flash performs surprisingly well compared to a slower MXFP4 version, showing no struggle in coding tasks and handling 200k+ context lengths. The post asks if any specific tasks are more fragile to quantization.

Original post →

More from Models

Models channel →