B300 spot prices hit $2.2M per unit in China, 3x premium pushes domestic chips into the推理 sweet spot
aigclink · x · 2026-09-11
A Chinese operator reports B300 spot prices in China reaching 16M RMB per unit vs under 5M RMB overseas — a 3x premium.
Key takes:
- The premium triples per-token inference cost on B300 in China, pushing it to parity with domestic chips
- Domestic accelerators become the optimal choice for inference operators not because they improved, but because of Nvidia's distorted pricing — which will drive domestic chip shipments and iteration
- Author predicts the balance shifts to a price war after Dec 2026, with supply-demand normalizing by mid-2027
- Context: national cluster vacancy rates exceeded 50% before Oct last year; the author's own clusters ran 30% vacant
Bottom line: the good times last about a year.
More from Infra
- 176 KB C Program Runs 2.78T-Param Kimi K3 on a Single CPU with 8.24 GB RAM — techNmak · 2026-09-11
- Blackstone's Biggest AI Bet Is Compute, Backing Deals with Google, Nvidia, Anthropic — abhiadesai · 2026-09-11
- Frontier models now independently reach for speculative decoding and kernel optimization on InferenceBench — maksym_andr · 2026-09-11
- A Beginner-Friendly Guide to Budget Multi-GPU Local LLM Setups — lblblllb · 2026-09-11
- Chinese Nvidia challenger Enflame jumps 179% in Shanghai debut, raises $910M — pstAsiatech · 2026-09-11
- Qwen3.8 Flash Next hits 49 tok/s locally on 2x RTX 3090 with FlashNext llama.cpp fork — whiteh4cker · 2026-09-11