Self-hosting AI: what rigs do local LLM runners actually use?
Fun_Kangaroo512 · reddit · 2026-09-06
A Reddit thread asks self-hosters to share their local AI setups. The typical architecture described: a high-RAM desktop or server runs the model around the clock and is accessed remotely, so the client device can stay lightweight — even a MacBook Air or a smartphone.
The poster asks commenters to list full specs, making the thread a useful reference for anyone planning to build a local LLM stack (RAM/VRAM sizing, model choices, remote access patterns).
More from Infra
- Wall Street veteran: the world's biggest data center consumes water like just 3 golf courses — rohanpaul_ai · 2026-09-06
- Japan says $550B U.S. investment pact advancing, AI and semiconductors to play major role — Polymarket · 2026-09-06
- QuixiAI open-sources SlimServe, the inference stack behind the 8x3090 Qwen deployment — QuixiAI · 2026-09-06
- Serving Qwen3.8-Flash-Next at 262K context on 8 RTX 3090s hits 1,200 tok/s — QuixiAI · 2026-09-06
- T-Glass shortage worsens: Kinsus losing 10-15% of monthly ABF revenue, 25% capacity expansion planned for 2027 — zephyr_z9 · 2026-09-06
- Bump-less 3D stacking goes practical: Intel Diamond Rapids first, AMD Zen rumored next — bookwormengr · 2026-09-06