llama.cpp runs faster on E-cores than P-cores in a GPU-offloaded MoE test

dir3ctly · reddit · 2026-07-25

Original post →

More from Infra

Infra channel →