Coding agents on 8GB VRAM: little-coder vs barebones just Pi for small models
Botoni · reddit · 2026-09-07
A developer running Hermes as a coding-agent harness finds its context footprint too heavy for a laptop with an 8GB NVIDIA GPU and 40GB RAM, and is weighing lighter options for models like qwen3.6 35b-a3b.
Candidates: little-coder (and smallcoder), which promise a small context and Pi-based extensions tailored for small models, versus just running barebones just Pi with its minimal system prompt and tools. The post asks for community tests and experience on whether the tailored extensions justify the overhead.
More from coding & agent
- LLM-powered revival of Put-That-There brings speech and gesture window control to XR — twi_mar · 2026-09-07
- Anthropic Claude Code engineer on internal AI coding practices and autonomy share — trq212 · 2026-09-07
- Leak: OpenAI to unveil Managed Agents at DevDay 2026 with hosted or self-hosted deploy — testingcatalog · 2026-09-07
- Harness-only changes lift deepagents-cli from 52.8% to 66.5% on Terminal-Bench 2.0 — Gauri_the_great · 2026-09-07
- Chops: open-source macOS app to manage AI agent skills across Claude Code, Cursor, Codex — tom_doerr · 2026-09-07
- LLMs vs classic CV: 200-line pipelines beat token-burning agents on rote tasks — mervenoyann · 2026-09-07