llama.cpp PR claims 15% faster ROCm prompt processing and fixes Q2_K bug

Betadoggo_ · reddit · 2026-07-21

A new llama.cpp PR claims about a 15% boost in prompt processing on ROCm and fixes a bug that made Q2K up to 28× faster.

Original post →

More from Infra

Infra channel →