Apple reportedly building enterprise AI server with its own M8 Ultra chips, launching no earlier than 2029
The Decoder · rss · 2026-09-17
According to The Information, Apple is developing an enterprise-grade AI server with two or four of its own M8 Ultra chips, targeting the AI inference market with a launch no earlier than 2029. Apple is also weighing Nvidia's NVLink Fusion technology for chip interconnects. The project could benefit from OpenAI and Anthropic already buying Mac hardware in bulk for AI workloads, signaling Apple's move from consumer devices into enterprise inference infrastructure.
More from Infra
- MLX-Serve 26.9.3 ships: Qwen Flash Next tops 100 tok/s on M4/M5 Max Macs — TheMoonMidas · 2026-09-17
- Running a 14B Model on 16GB RAM: 'My PC Is a Toaster Now' — Aggravating_Site381 · 2026-09-17
- fal engineering head: we'll never pre-train, inference compute is the real moat — jfischoff · 2026-09-17
- Meta extends FlashAttention-4 with MXFP8 on Blackwell, hitting 2.85 PFLOP/s forward — PyTorch · 2026-09-17
- AWS shows NVRx fault-tolerant FSDP training on EKS: sync checkpointing ate up to 40% of wall time — AWS ML Blog · 2026-09-17
- Qwen 3.8 27B NVFP4 benchmarked on 2xV100 with ~400k context — jjusko20 · 2026-09-17