Compute Benchmarks: RTX 5090 vs 6000 PRO
panchovix · reddit · 2026-07-12
The author provides a detailed performance comparison of the RTX 5090, 6000 PRO MaxQ, and 6000 PRO WS/SE, covering image generation and local LLM inference.
They performed a shunt mod on the 6000 PRO MaxQ, combined with water cooling, pushing the power limit to around 600W, with temperatures ranging between 45°C and 60°C. A rented 6000 PRO WS from runpod was used as a baseline.
Test software and setups include:
- Various PyTorch versions
- SageAttention 2.1
- Forge neo
- RTX upscaling extensions and additional sampler extensions
- torch compile using max autotune without cudagraphs
Image generation tests were run with fixed samplers, steps, and prompts. For LLMs, llama.cpp was used with partial model offloading to the CPU, bottlenecking the main GPU to test local inference performance. Overall, the post highlights how different GPUs, power limits, and cooling solutions impact AI compute throughput.
Related event: Performance Comparison: RTX 5090 vs RTX 6000 PRO(2 posts)→
More from Infra
- Nothing phone mockup turns a film joke into a modular design meme — ZeYanjie · 2026-07-22
- Actual Computer says its inference stack is tuned for Nvidia’s consumer Blackwell lineup — markjeffrey · 2026-07-22
- Ben Bajarin says CPU demand is still being badly underestimated — BenBajarin · 2026-07-22
- An energy model says the U.S. could run short of natural gas starting in 2028 — churchkey · 2026-07-22
- Devin adds e2b sandboxes for remote agent execution — badphilosopher · 2026-07-22
- Arbitrum fee simulation shows higher gas capacity but lower L2 revenue under ArbOS61 — tomwanhh · 2026-07-22