New local AI app hits 164 tokens/sec on iPad, tests to follow
TheOyinbooke · x · 2026-09-26
A developer teased an unreleased local AI app running at 164 tokens/sec on an iPad, promising more benchmarks and opening sign-ups for testers. If confirmed, that on-device speed would make local LLMs notably more usable on mobile, though details on the model, quantization, and hardware are not yet disclosed.
More from Infra
- SpaceX's Mid-South AI clusters span 2.5M sq ft with millions of GPUs and over 2GW of compute — chaitu · 2026-09-26
- Prime Intellect posts 30 open roles, betting full-stack on open AGI infrastructure — willcb · 2026-09-26
- IIT Delhi researchers unveil India's first indigenous micro-GPU — rvp · 2026-09-26
- Prime Intellect Maps the Emerging AGI Stack: RLaaS, Evals, Inference, Sandboxes and More — willcb · 2026-09-26
- Fireworks x HUD Release RL Training Cookbook: Define Task Once, Train and Evaluate LoRA — sophiamyang · 2026-09-26
- Inside the First Tokenomicon: Amsterdam Event on LLM Token Economics — Bartaseth · 2026-09-26