An $8 ESP32-S3 now runs a 28.9M-parameter language model offline
brianrkelly · x · 2026-07-24
- A developer has forced a 28.9-million-parameter language model onto an ESP32-S3 microcontroller that costs about $8.
- The model runs fully offline, generates short stories at roughly 9.5 tokens per second, and uses power comparable to a small LED.
- The poster frames it as a dramatic expansion of what “local AI” can mean on tiny hardware.
- He also suggests a network of many such devices could specialize on different domains and report back to a master node.
Related event: Developer Runs 28.9M Parameter Model Offline on $8 ESP32-S3(2 posts)→
More from Infra
- AI governance chatter is cooling as automation and data plumbing keep rising — YvesMulkers · 2026-07-24
- An $8 ESP32-S3 now runs a 28.9M-parameter model fully offline — brianrkelly · 2026-07-24
- NVIDIA releases a strong multilingual 1B embedding model for retrieval — tomaarsen · 2026-07-24
- AI ETF thread says markets have already priced in Qualcomm, AMD and other August events — alejandroll10 · 2026-07-24
- Lidl starts rolling out Wero to undercut Visa and Mastercard in Europe — MarvinTBaumann · 2026-07-24
- Can AI-Toolkit train a Krea 2 LoRA on an RTX 5080 with 16 GB of VRAM? — NexusFred · 2026-07-24