DeepSeek V4 Flash Local Quantization Benchmark on SlopCodeBench

corruptbytes · reddit · 2026-08-09

A developer shared local quantization benchmark results for DeepSeek V4 Flash on SlopCodeBench. Running with antirez q2-q4 imatrix quantization on a MacBook M5 Max, it is much slower than the hosted API, but switching the harness to pi 0.84.0 helped recover some of the intelligence lost to quantization.

Key Results (Strict category):

The test indicates that proper harness configuration is crucial for maximizing the performance of locally quantized models.

Original post →

More from Infra

Infra channel →