LARRI: open-source CLI that rents a GPU and tunnels open models to a local endpoint
gethackteam · x · 2026-09-09
LARRI is a simple open-source CLI that rents a GPU, spins up any open-weights model on it (Qwen, GLM, Devstral, etc.), and tunnels it to a fixed local endpoint—with commands to spin instances up and down. It solves the pain of maintaining your own inference infrastructure while still running local models.
Related event: LARRI: Open-Source CLI Spins Up Open Models on Rented GPUs with One Command(2 posts)→
More from coding & agent
- Semantic Cache Verification Measured: 30-40ms on the Hot Path vs 1.7s for an LLM Judge — Reasonable_Royal_621 · 2026-09-09
- HyperFrames hits #1 on GitHub as Claude Code-made launch video clocks 3M views in 2 days — toolstelegraph · 2026-09-09
- AI assistants may need their own email addresses as users balk at full inbox access — _AustinCalvert_ · 2026-09-09
- Dimillian's dev tool adds worktree creation on any branch after user requests — Dimillian · 2026-09-09
- Dev built a custom procedural drawing engine via Astra — all Evergrow art is drawn in code — Dimillian · 2026-09-09
- Agents researching agents: Prime Intellect's open stack eases RL environments — seanwbren · 2026-09-09