4x3090 Benchmarks: 27B Q8KXL at 32GB Beats Next IQ3 at 82GB on Both Speed and Security Tasks

Repulsive_Initial308 · reddit · 2026-08-28

A user running llama.cpp on 4x RTX 3090s shared initial tests of the 3.8 Next IQ3 quantization, calling it "meh." Their existing mix of Q4/Q8/BF16 already feels near-perfect. Early results show 3.8 27B Q8KXL at 32GB of weights is vastly superior to 3.8 Next IQ3 at 82GB — worse in security vulnerability assessment AND slower. He notes it's early days, so things may change.

Original post →

More from Infra

Infra channel →