Hugging Face hosts 568 plug-and-play compute kernels, filterable by 200+ hardware platforms

ariG23498 · x · 2026-10-07

Hugging Face's Kernels hub now lists 568 compute kernels that can be browsed like models and datasets, filtered by hardware (H100, A100, Apple M4, WebGPU, and 200+ more) and plugged directly into your model. Popular entries include flash-attn3 (28.9k), flash-attn2, deep-gemm, and ggml quantization/attention kernels across CUDA, ROCm, and Metal.

Original post →

More from coding & agent

coding & agent channel →