Intel GPU Experience: Running Qwen3.8 27B with 116k Context on a Budget

Accomplished_Yard636 · reddit · 2026-08-20

The author shared their experience running Qwen3.8-27B BF16 locally on a $3K 64GB Intel GPU. It achieves 16 tps with 116k FP8 token context, which is sufficient for coding needs. The author praised the out-of-the-box suspend functionality, noting it is more convenient than Nvidia's solutions.

Original post →

More from Infra

Infra channel →