llama.cpp Vulkan benchmarks: mining card P102-100 tops 76-GPU price-performance ranking

tabletuser_blogspot · reddit · 2026-10-09

A Reddit user compiled llama.cpp Vulkan benchmark data (76 GPUs, Llama 2 7B Q40) against 2026 secondhand prices into a local-LLM price-performance ranking. Nvidia's $40-50 P102-100 mining card tops the list (1.3 tok/s per dollar decode) thanks to its 320-bit bus; AMD Instinct MI50 and Radeon VII win on decode cost per token via 1TB/s HBM2. Wide-bus older flagships like GTX 1080 Ti and RTX 2080 Ti often out-decode modern cards under $300, while modern mid-range cards are bottlenecked by narrow 128/192-bit buses.

Original post →

More from Infra

Infra channel →