iPhone chip's 50% memory bandwidth jump over A19 Pro matters more for local AI than 2nm

HankYeomans · x · 2026-09-10

Quoting PrinceCanuma, HankYeomans argues that everyone will cite the new iPhone's 2nm process, but the number that matters for local AI is memory bandwidth: roughly 50% more than the A19 Pro.

LLM decoding is bandwidth-bound — tokens/sec comes from moving weights faster, not from faster cores. He calls it the biggest single-generation bandwidth jump he's seen in an iPhone, with far more impact on on-device inference than the process node.

Related event: A20 Pro Memory Bandwidth Up 50%, Key for On-Device AI(2 posts)→

Original post →

More from Infra

Infra channel →