XeBoostLM: native C++ local LLMs on Intel NPUs and iGPUs, zero Python

Spiritual-Ad-5916 · reddit · 2026-09-17

A developer released XeBoostLM v0.1.2, a pure C++ CLI built on OpenVINO GenAI for running local LLMs on Intel Core Ultra NPUs, Arc iGPUs, and CPUs with zero Python overhead. It offers manual or HYBRID hardware routing, an OpenAI-compatible SSE streaming server that plugs into Open WebUI and LibreChat, INT4/INT8 model pulls (Qwen 2.5, Phi-3.5, Llama 3.2, DeepSeek R1), and a terminal REPL. Open source, feedback welcome.

Original post →

More from Infra

Infra channel →