Apple reportedly plans M8 Ultra AI inference servers for enterprises, backed by new CEO Ternus
APPSO · wechat · 2026-09-28
- Apple is reportedly discussing a return to the enterprise server market (abandoned since 2011's Xserve) with machines packing two to four M8 Ultra chips, possibly linked via NVIDIA's NVLink Fusion, to run pre-trained models for inference — not training. New CEO John Ternus reportedly backs the project.
- The piece frames this via the local-AI boom: OpenClaw-style always-on agents made high-memory Mac minis scarce, while 175,000 misconfigured Ollama instances exposed on the public internet show the risks of DIY local AI. Enterprises want private inference without building NVIDIA-based stacks from scratch.
- Apple's unified memory (up to 512GB, 1.2TB/s on M5 Ultra) makes it well suited to turnkey inference boxes; it sells machines, not cloud compute.
- Ternus, Apple's first hardware-background CEO in 30 years, is doubling down on deciding where compute happens: on-device first, then Mac, then private servers, with cloud as overflow — while Apple leases Gemini for Siri at $1B/year.
More from Companies & People
- Personal vs office agent gateways: the real problem is rules that don't travel between them — sujingshen · 2026-09-28
- MIT-licensed AI engineering curriculum with 523 hand-built-algorithm lessons ships as free EPUB/PDF books — SeveralSeat2176 · 2026-09-28
- Musk: Starship was designed without AI, likely the last big pre-AI feat — XFreeze · 2026-09-28
- EA organizations slammed for 'unhinged' hiring: 10-20 work trials over months — AaronBergman18 · 2026-09-28
- Founder brings back the wolf, CEO kills it: a VC's parable on the founder-CEO trust relationship — yongqianme · 2026-09-28
- 'Why did Mistral stop trying to build frontier models?' — community laments pivot — burny_tech · 2026-09-28