Seeking Real-World Benchmarks: R9 7900XTX Running Qwen2.5-72B
BillyQ · reddit · 2026-08-24
The user plans to buy a PowerColor R9700 (7900XTX) to run Qwen2.5-72B (Q4KXL) via llama.cpp/Vulkan and is seeking real-world token/sec numbers at 64K+ context length under real workloads. Despite AMD's claim of 51.8 tok/s, benchmarks on the NVIDIA 5090 show significant performance degradation for Qwen2.5 as context fills. The user is also inquiring about the stability of MTP speculative decoding to avoid OOMs or garbage output.
More from Embodied
- Autonomous 5v5 humanoid soccer kicks off; fall recovery improves but possession still hard — rohanpaul_ai · 2026-08-24
- Giant rideable robot dog revealed with mind-blowing demo — thetripathi58 · 2026-08-24
- Chinese robots now outrunning the best human athletes — thetripathi58 · 2026-08-24
- Robotics stack consists of two loops: millisecond control and monthly industrial — demian_ai · 2026-08-24
- From Jurassic Park's Dinosaur Input Device to Hand-Tracked Blender in Vision Pro — bilawalsidhu · 2026-08-24
- The Harness Is Becoming the Operating System for Physical Intelligence — eigenron · 2026-08-24