DeepSeek V4 Flash Local Quantization Benchmark on SlopCodeBench
corruptbytes · reddit · 2026-08-09
A developer shared local quantization benchmark results for DeepSeek V4 Flash on SlopCodeBench. Running with antirez q2-q4 imatrix quantization on a MacBook M5 Max, it is much slower than the hosted API, but switching the harness to pi 0.84.0 helped recover some of the intelligence lost to quantization.
Key Results (Strict category):
- Local quant (run B): 5/17 (29.4%)
- Opus 5 (API): 4/17 (23.5%)
- DeepSeek V4 Flash (API): 3/17 (17.6%)
The test indicates that proper harness configuration is crucial for maximizing the performance of locally quantized models.
More from Infra
- Microsoft Analyzes 13.5M Copilot Sessions: Why Agent Scheduling Differs from Chat — rohanpaul_ai · 2026-08-09
- Beef vs. Data Centers: The Double Standard in AI Water Consumption Debates — KrustyKrabFormula_ · 2026-08-09
- Replace ChatGPT Plus with Local Models: A 5-Step Guide — Aiden_Tech_Ai · 2026-08-09
- Nvidia's Rubin Ultra Shifts from HBM to Optical Interconnects, Altering Market Dynamics — zephyr_z9 · 2026-08-09
- Running SD Natively on Android: SDXL Takes 20 Minutes on a Phone — Silent-Paramedic4063 · 2026-08-09
- LFM 2.6B Hits 260 Tokens/s on RTX 3090: A Dev's Hands-On Review — Borkato · 2026-08-09