Apple M5 Ultra decoding speed crushes DGX, NVIDIA's moat under multi-pronged attack

XFreeze · x · 2026-08-26

The post argues that NVIDIA's long-standing AI compute moat is being attacked from two directions in a single day: cloud inference and local/edge inference. Cited benchmarks show that a 256GB Mac Studio M5 Ultra achieves a decode speed of 1200GB/s, nearly 4.4 times faster than a dual DGX Spark (273GB/s), while prefill speeds are almost tied. Given the similar price point ($10,000), the author deems the Mac Studio a better value, validating Elon Musk's prediction that digital outputs are easily replicated and outperformed by AI.

Related event: M5 Ultra Mac Studio vs NVIDIA RTX 6000 for Local AI(4 posts)→

Original post →

More from Infra

Infra channel →