Running Qwen 3.8 27B on RTX 3090: Configuration Guide
cezarducatti · reddit · 2026-08-18
A user shares their configuration for running the Qwen 3.8 27B model on an RTX 3090 (24GB VRAM). Switching from Q4XL (used in v3.6) to Q3XL with reasoning mode enabled for v3.8, the author reports excellent performance, including successfully refactoring a 4,000+ line Python script. The post details cache settings, a 160k context window, and the full command-line arguments used.
More from Infra
- Cursor releases Origin: a Git version optimized for AI agents — nicolascraske · 2026-08-18
- $10T in AI datacenter capex blocked: GPUs need 230GW, US grid delivers ~100GW — PeterDiamandis · 2026-08-18
- Power Grid Bottlenecks Stall $10T AI Data Center Buildout — PeterDiamandis · 2026-08-18
- Discussion: Running bots locally on Linux VMs — Daniel_Farinax · 2026-08-18
- PotatoMesh: Federated Dashboard for Visualizing LoRa Node Positions — tom_doerr · 2026-08-18
- llama.cpp adaptive MTP PR speeds up code generation by up to 100% — Look_0ver_There · 2026-08-18