Which local LLM should run a long-lived agent on 36GB of VRAM?
hovikyan · reddit · 2026-07-28
This is a repost of the same question about choosing a local LLM for an agent that can keep working on software projects and non-coding tasks on a machine with roughly 36GB of VRAM. The candidates are the same set of local models: Qwen3.6 27B, Qwen3.6 35B A3B, Gemma 4 31B, GLM 4.7 Flash, and Llama 3.3 70B.
Related event: Selecting Local Agent Models for 36GB VRAM(2 posts)→
More from coding & agent
- Waddle Labs launches as “Claude Code for robots” with API-driven robot control — ycombinator · 2026-07-28
- AI won’t shrink software demand—it may unlock a much larger services market — joecole · 2026-07-28
- A practical cloud-agent stack: Devin, Cursor, Open-Inspect, Infisical and more — vinvan · 2026-07-28
- Stripe upgrades Directory with instant listings and MCP search for agents — jeff_weinstein · 2026-07-28
- AIWayfinder Launches DCA v0.6.0 with Zero-Intervention Self-Healing Daemon — templecrash · 2026-07-28
- Council 1.2 adds blind multi-model review and a guest seat for any external answer — ahumanbeingmars · 2026-07-28