Local LLM crowd shares deployment speeds but never mentions evals
_nateraw · x · 2026-09-03
Developer nateraw complains about a common pattern in the local LLM space: people share deployment speeds while saying nothing about evals — fast serving says little about model quality.
More from Infra
- WSL nested virtualization is coming: team implements it after user's PoC PR — unixterminal · 2026-09-03
- Vercel Fluid Compute Now Runs 15M Builds Daily, Unifying Functions, Sandboxes and Builds on One System — soleio · 2026-09-03
- Baseten, NVIDIA Dynamo and SGLang to host SF meetup on RL post-training infrastructure — BanghuaZ · 2026-09-03
- Running Code OSS in a Cloud Run instance for a few bucks a month — rseroter · 2026-09-03
- Matthew Berman: data centers might get solved — Matthew Berman · 2026-09-03
- One startup's token usage jumped 720x in two months, from 1B per month to 1B per hour — MartinGTobias · 2026-09-03