Benchmarked: Apple Core AI vs MLX for on-device LLM speed on iPhone and Mac
HankYeomans · x · 2026-09-08
A new hands-on benchmark compares Apple's Core AI framework against the open-source MLX stack, running the same model on both iPhone and Mac to see which delivers faster on-device inference. The piece focuses on real speed differences across the two devices and frameworks.
More from Infra
- Actian's new vector DB claims 22x Qdrant speedup, but its docs look suspiciously like Qdrant's — qdrant_engine · 2026-09-08
- 8GB VRAM runs Flux and Wan 2.2 locally fine: three wrong settings, not the GPU, were the bottleneck — leonbuilds · 2026-09-08
- China effectively leads the humanoid robot supply chain, and Optimus relies on it — JOBhakdi · 2026-09-08
- Tiered KV-cache offloading for self-hosted LLM inference: GPU to RAM to NVMe to S3 — Responsible-You9024 · 2026-09-08
- D-Wave Finalizes Agreement with US Commerce Dept for Up to $100M in CHIPS Act Funding — ceciletamura · 2026-09-08
- Export Controls Working? H200 Sells for 280 and B300 for 450 Overseas — teortaxesTex · 2026-09-08