Open Weights Don't Mean Locally Runnable
Mountain_Patience231 · reddit · 2026-07-17
This post vents frustrations about the recent wave of "open weight" LLM releases. While models like GLM-5.2 boast impressive parameter counts, context lengths, and licenses, they aren't actually "locally runnable" for the average user.
The author's main points are:
- 700B-scale MoE models are nearly impossible to fit into consumer hardware, even when quantized
- These releases feel more like marketing for cloud/API models that most people can "see but can't run"
- The community's previous focus on self-hosting, tinkering with llama.cpp, and quantization is being diluted by these ultra-large model discussions
While not denying the value of open weights, the author argues that for everyday users, these giant models are practically no different from closed-source alternatives.
Related event: Growing Open-Source Weights Spark Local Deployment Concerns(3 posts)→
More from Models
- Users say GPT-5.6 Ultra feels like extra token burn with little visible gain — CtrlAltDwayne · 2026-07-21
- Early Gemini 3.6 Flash outputs look fast but weak on frontend and spatial reasoning — max_paperclips · 2026-07-21
- Anthropic removes Fable’s access deadline, but users say it was nerfed — oykun · 2026-07-21
- Kimi K3 retakes first place on DesignArena’s frontend web app benchmark — rohanpaul_ai · 2026-07-21
- Last Week in AI roundup covers Claude Sonnet 5, LongCat 2.0, and new agent benchmarks — Last Week in AI · 2026-07-21
- Rumor claims GPT-6 could arrive in August — iruletheworldmo · 2026-07-21