Hugging Face releases 207 open-source WebGPU kernels for browser inference
nicodotdev · x · 2026-09-01
Hugging Face ships @huggingface/kernels on npm, providing 207 open-source WebGPU kernels loaded straight from the Hub. This addresses the bottom layer of the fast browser inference stack. They also launched Fleet, a tool to benchmark your GPU in the browser and get a card for your device.
Related event: Hugging Face open-sources 207 WebGPU kernels for faster browser inference(2 posts)→
More from Infra
- GLM-5.3-Flash beats DeepSeek-V4-Flash for writing and vision on 2× DGX Spark — kuhunaxeyive · 2026-09-03
- Jeff Dean's 2007 slide on hardware failures inside a datacenter resurfaces — SumitGup · 2026-09-03
- Google unveils 8th-gen TPU at Hot Chips: two chips per year, split inference and training designs — firstadopter · 2026-09-03
- Data center water consumption is a myth, says engineer: modern builds use closed-circuit cooling — GlenBradley · 2026-09-03
- Inference Engineering Is Just a Recipe: vLLM/SGLang, Replicas, Cache-Aware Routing — GabGarrett · 2026-09-03
- Databricks pitches agent-native data infrastructure, Lakebase Postgres at VLDB 2026 — matei_zaharia · 2026-09-03