Halv desktop workspace cuts coding-agent token use by 51% and lifts solve rate on SWE-rebench
EnslavedFish · reddit · 2026-09-06
A developer shipped Halv, a desktop app that runs Claude Code alongside other coding agents on existing subscriptions, with chat, split terminals, saved sessions, and a live savings meter. Under the hood it compresses context, filters noisy command output, and maintains a code index to help agents find what they need.
A self-run benchmark of Codex on 20 paired SWE-rebench tasks showed:
- 51.1% fewer tokens per correct answer
- 30.2% fewer total tokens
- 10/20 tasks solved vs 7/20 without Halv
All 40 run records and verifier results are published on halv.ai.
More from coding & agent
- Cognition rumored to launch a new model soon, per AI insider hunch — realsohamparekh · 2026-09-06
- Redditor builds free AskSary creative studio with game engine GPT can play and patch live — Beneficial-Cow-7408 · 2026-09-06
- Do the fun terminal work yourself, let Claude handle the boring chores — 4310sy · 2026-09-06
- Heavy AI user's cost ledger: 100+ agents a day, $70 of DeepSeek in two days — simmon_charlie · 2026-09-06
- First untrained agent run on local Qwen 3.8 Flash, no skills configured — jasonkneen · 2026-09-06
- Is a schema-aware memory graph 'overfitting'? Dev asks for the cleanest leakage test — chaachans · 2026-09-06