Which local LLM is best for a 36GB VRAM coding agent setup?

hovikyan · reddit · 2026-07-28

The poster has about 36GB of VRAM and wants a local LLM to drive an agent like Hermes or OpenCode on software projects and other personal tasks. They ask which of several locally runnable candidates—Qwen3.6 27B, Qwen3.6 35B A3B, Gemma 4 31B, GLM 4.7 Flash, or Llama 3.3 70B—would be best for that setup and use case.

Related event: Selecting Local Agent Models for 36GB VRAM(2 posts)→

Original post →

More from coding & agent

coding & agent channel →