Qualcomm's next-gen Hexagon NPU runs 30B MoE models with 32K context on-device
lee_stott · x · 2026-09-11
Qualcomm detailed its next-gen Hexagon NPU for Snapdragon devices: a new Element Accelerator for transformers, +50% shared memory, 32K context window, the ability to run 30B-parameter MoE models (3B active per token), and +50% INT4 prefill speed — a big step for agentic on-device AI.
Related event: Qualcomm's next Hexagon NPU to run 30B MoE models on-device(2 posts)→
More from Embodied
- Replit Agent in a robot builds and publishes websites autonomously via MCP — amasad · 2026-09-11
- TARS Robotics unveils embodied foundation model AWE: 15+ tasks, one model, zero retraining — heyshrutimishra · 2026-09-11
- Working with an ESP32 device using Copilot CLI and Astra — DanWahlin · 2026-09-11
- VATIX: open-source driving world model trained on 5,500 hours of real-world footage — abursuc · 2026-09-11
- Reachy Mini robot gains spotlight after joining NVIDIA's portfolio — RachelVT42 · 2026-09-11
- PhyFilter: physics-informed filter lets sim-trained robots traverse real terrain without extra data — 机器之心 · 2026-09-11