RTX 3090 Inference Power Test: Performance Not Linear with Power Draw
haluxa · reddit · 2026-08-30
A user conducted an informal power consumption test for inference on an RTX 3090 using LMStudio and the Qwen3.8-27B-Q4KM model (20GB VRAM). Results indicate that inference speed is not linearly correlated with power draw, suggesting an optimal power range exists. The previously assumed 220W sweet spot may vary depending on the inference engine and model.
More from Infra
- Qwen 350K Context Tested on M5 Max: Performance and Quality — Artistic_Okra7288 · 2026-08-30
- Azure Linux 4.0 Desktop Concept: PowerShell, Edge, and Copilot Pre-installed — unixterminal · 2026-08-30
- Jensen Huang: Built GPU tech first, found endless problems from graphics to molecular dynamics — r0ck3t23 · 2026-08-30
- How to build an LLM inference engine from scratch: 5-layer architecture — glenbeer · 2026-08-30
- Huaqin expects super node revenue to exceed 10B RMB in 2H 2026 — zephyr_z9 · 2026-08-30
- Nvidia is generating $1 billion a day, a business scale deemed absurd years ago — shauntrennery · 2026-08-30