Running Qwen2.5-14B Locally on RTX 5060 Ti 16GB: Hits 44 t/s Generation Speed

Primary_Olive_5444 · reddit · 2026-08-13

A developer shared performance metrics for running Qwen2.5-14B-Instruct (Q4KM quantization) locally on an RTX 5060 Ti 16GB:

Original post →

More from Infra

Infra channel →