LexiPanel: open-source AIO web panel for llama.cpp inference servers
W61k3r · reddit · 2026-09-20
A developer open-sourced LexiPanel, a lightweight web-based management panel for llama.cpp inference servers, originally built to test Qwen3.8-27B.
Features include real-time monitoring of server status, GPU utilization and memory tracking, model loading/unloading with multi-model support, configuration management (context size, batch size, threads), optimization profiles, a web terminal via ttyd, and instance management.
Install with ./install.sh and run with ./start.sh; the panel serves on port 8090. Requires Python 3.8+, llama.cpp, Caddy for TLS, and ttyd. MIT licensed.
More from Infra
- jevcache memoizes model decisions by (model, schema, state) — repeat calls cost $0 and return in ~0ms — JiliJeanlouis · 2026-09-20
- Brain Runs on 20 Watts: Can Neuromorphic Computing Make AI Less Power-Hungry? — burny_tech · 2026-09-20
- Ternary 2-bit Bonsai-2-27B GGUF lands on Hugging Face trending — dealignai · 2026-09-20
- Pedro Domingos: ASML Is a Single Point of Failure for the AI Supply Chain — pmddomingos · 2026-09-20
- On-device model Cactus Needle trends on Hugging Face with tool-calling support — Cactus-Compute · 2026-09-20
- Qwen3.8 Flash Next on one RTX 5090 hits 50 t/s decode via FreeToken expert caching — dir3ctly · 2026-09-20