Andrew Chen: strong LLMs are far from running on phones, on-device AI faces bandwidth, heat and model-size hurdles
andrewchen · x · 2026-09-21
Andrew Chen argues that strong LLMs remain far from native on-device execution: slow memory bandwidth, only highly quantized MoE models fit, and power/heat remain issues. Next-gen mobile NPUs target only modestly sized LLMs.
He sees major mobile UX opportunities — notifications, typing, inboxes, calendars — that would benefit from fast, cheap AI decision models, and floats the idea of baking an older-but-useful model directly into phone hardware, wondering if Jevons-like dynamics could accelerate the shift.
More from Embodied
- Human vs Robot Fight at San Francisco's REK Event Hit 'Completely Different', Says Attendee — BLUECOW009 · 2026-09-21
- DIY robot U-BOT's maiden voyage: stuck on a tiny stick, grass proves tougher than expected — _Stocko_ · 2026-09-21
- First-ever Human vs Terminator robot fight pits Frankie LaPenna against a bot — BLUECOW009 · 2026-09-21
- Astra shows any 3D/4D prior can be distilled into VLMs, a new embodied AI paradigm — mariyaivasileva · 2026-09-21
- Legless autonomous food robot sparks debate: fixed-purpose machines land before humanoid cooks — mrjonfinger · 2026-09-21
- Closed-loop spatial understanding from just two wrist cameras, no gripper feedback — ChongZzZhang · 2026-09-21