Apple Approves MLX LLM Serving App in Under 10 Hours, Stunning 19-Year iOS Dev
alexcovo_eth · x · 2026-09-21
Developer @ddalcu's MLX-Serve on iPhone app — which runs LiquidAI models with local on-device inference — was approved by Apple in under 10 hours. The developer says this never happened in 19 years of iOS development, publicly thanking Apple and executive John Ternus.
He had previously joked that App Store review typically takes 3 months just to learn a rejection reason, so the lightning-fast approval is a striking contrast — possibly signaling a shift in Apple's stance toward on-device AI inference apps.
More from Infra
- vLLM ships Hybrid KV Cache Manager for mixed-attention model inference — TheZachMueller · 2026-09-22
- SGLang's hicache: use an L3 storage cache to keep KV cache alive across local model swaps — TheZachMueller · 2026-09-22
- Egypt's AI Ecosystem Hits Production Scale With $400M Data Center, 10x NVIDIA Learner Growth — nordicinst · 2026-09-22
- RTX Pro 6000 vs a used 3090 vs cloud rental: the LoRA training math — big-in-jap · 2026-09-21
- ComfyUI GPU rental showdown: Modal's 35s cold starts and free 1TiB beat RunPod — ronalder100 · 2026-09-21
- Meta partners with Arm on Arm AGI CPU, its first AI-era data center CPU — bookwormengr · 2026-09-21