Hugging Face open-sources 200+ WebGPU kernels for in-browser local AI
xenovatech · reddit · 2026-10-01
A developer has open-sourced what they call the world's fastest WebGPU kernel collection for local AI on Hugging Face, covering more than 200 common ML operations that run entirely locally in the browser via WebGPU. The author says they are working to upstream these optimizations into Transformers.js, ONNX Runtime Web, LiteRT.js and more.
- Kernels: huggingface.co/kernels?platform=webgpu
- Blog: huggingface.co/blog/webgpu-kernels
- Goal: faster in-browser local inference, less reliance on cloud compute
More from Infra
- RAG and RL tool calling are bringing CPUs back into ML training, possibly needing dedicated CPU nodes — StasBekman · 2026-10-01
- China's CXMT to spend $5.2b on DRAM expansion, favoring domestic equipment suppliers — pstAsiatech · 2026-10-01
- Cloudflare launches agent-themed batch: pay-per-use gateway, AutoRouter, 6x faster containers — threepointone · 2026-10-01
- Distributed compute market adds 10 B300 nodes, rentable from a single node — markjeffrey · 2026-10-01
- 27B at Q5 with full 131k context on one 24GB RTX 3090, 13-17% faster — bjivanovich · 2026-10-01
- netkit paper: container network namespaces cost up to 31% throughput on Linux — tianyin_xu · 2026-10-01