Dev ships ik_llama.cpp quants for DeepSeek V4 Flash, beats Atomic in paired KLD tests

KeinNiemand · reddit · 2026-08-28

KeinNiemand published a full ikllama.cpp GGUF ladder for DeepSeek V4 Flash 0731 on Hugging Face, from a 65.1 GB XSIQ1KT up to a 149.1 GB IQ4KSS.

Quant details

Paired quality testing

Using AtomicChat's published lossless BF16 reference logits and WikiText-2 tokens under an identical local setup, the author compared PPL, mean KLD, RMS delta-p and top-1 agreement:

Takeaways

Original post →

More from Infra

Infra channel →