Compute Benchmarks: RTX 5090 vs 6000 PRO

panchovix · reddit · 2026-07-12

The author provides a detailed performance comparison of the RTX 5090, 6000 PRO MaxQ, and 6000 PRO WS/SE, covering image generation and local LLM inference.

They performed a shunt mod on the 6000 PRO MaxQ, combined with water cooling, pushing the power limit to around 600W, with temperatures ranging between 45°C and 60°C. A rented 6000 PRO WS from runpod was used as a baseline.

Test software and setups include:

Image generation tests were run with fixed samplers, steps, and prompts. For LLMs, llama.cpp was used with partial model offloading to the CPU, bottlenecking the main GPU to test local inference performance. Overall, the post highlights how different GPUs, power limits, and cooling solutions impact AI compute throughput.

Related event: Performance Comparison: RTX 5090 vs RTX 6000 PRO(2 posts)→

Original post →

More from Infra

Infra channel →