AMD MI355X beats NVIDIA B300 on tokens-per-dollar in AgentX
AccBalanced · x · 2026-09-07
SemiAnalysis reports that an AMD MI355X submission with vLLM and LMCache beats the NVIDIA B300 on total tokens per dollar TCO at lower interactivity ranges on AgentX.
A viral quote-post jokes that markets need six months to reprice the "NVIDIA is safe" slide now that cheaper tokens are on the table.
More from Infra
- M1 64GB Mac is no match for local Astra: 'this thing is crawling' — natesiggard · 2026-09-07
- KV Cache Engineering for LLM Serving: 12 Techniques Explained With Trade-offs — AccBalanced · 2026-09-07
- SlimServe open-sources low-cost LLM serving for consumer GPUs — QuixiAI · 2026-09-07
- AI buildout debt hits $570B, now rivaling the entire US muni bond market's annual supply — ivan_bezdomny · 2026-09-07
- 928-Star Wiki Details Running Qwen3.5-397B and Kimi-K2.5 on NVLink-Free PCIe RTX 6000 — TheZachMueller · 2026-09-07
- Prediction: compute is moving from rack-scale to datacenter-scale as bottlenecks shift outward — AccBalanced · 2026-09-07