Baby's first PC: 3090 with local LLM inference
QuixiAI · x · 2026-08-20
User shared a setup for a child's first PC featuring a 48GB RTX 3090 running a local inference engine (SlimServe) with a quantized Qwen2.5-72B model.
Related event: Modded 48GB RTX 3090 Runs 27B Model Locally at High Speeds(4 posts)→
More from Infra
- Scholars note data center exploitation predicted by Decolonial AI movement — rajiinio · 2026-08-20
- Rust Compiler Contributor Sets Goal: Make rustc 10x Faster — mitsuhiko · 2026-08-20
- Paper: LLMs Beat Embeddings Slightly but Cost 1,431x More — Muennighoff · 2026-08-20
- Why hasn't AI pricing increased to its 'true cost' yet? — MacIntoic · 2026-08-20
- Obscura: Rust-based headless browser for AI agents — tom_doerr · 2026-08-20
- Unsloth AI to launch new kernels, predicting 300 tok/s on C8 — QuixiAI · 2026-08-20