browser-llm-fit: check if an AI model fits your browser before downloading weights
init0 · reddit · 2026-09-04
Developer hemanth released browser-llm-fit, an open-source tool tackling common WebGPU crashes when browser-run models exceed maxStorageBufferBindingSize or lack shader-f16 support. It probes client hardware limits and ranks browser-executable models before weights are downloaded. fit('model') tests one model (returning fit and estimated speed, e.g. 45-65 tokens/sec), fit() returns all models sorted by hardware fit. Live demo, GitHub repo, and npm package available.
More from Infra
- HPE Delivers Strong Q3 on AI Server Demand, Raises FY2026 Outlook — mattwbaker · 2026-09-04
- All Chromium Browsers Hit by Hard-to-Reproduce Data Loss Bug, Devs Say — uwukko · 2026-09-04
- Hundreds protest Scotland's datacentre boom, demanding pause on 20+ proposed projects — nordicinst · 2026-09-04
- Built a Dual RTX 6000 Pro Rig for Local DeepSeek — Warns Against Influencer Build Hype — HankYeomans · 2026-09-04
- Inference engines are an underexamined attack surface, self-hosting ops warned — JeremyCMorgan · 2026-09-04
- YC S26 Demo Day Next Week: Floating Data Centers, Diamond Semiconductors, Bio Computers — ycombinator · 2026-09-04