Dev open-sources a token compression CLI that cut his codex usage by 34%
Icy-General-9096 · reddit · 2026-09-30
An indie dev who builds web apps to replace SaaS subscriptions got tired of coding models burning through his usage resets, so he reasoned: if I'm paying per token, why not use fewer tokens? He built a free open-source CLI that sits as a local proxy in front of codex and compresses prompts and context, measuring a 34.3% token reduction — now left on by default, he can finish sessions without running out of credits.
- Repo: github.com/spenmcke/compress
- Privacy: local proxy with zero data retention, no queries stored
- Motivation: switched from Sol (buggy UI) to astra + opus (great but quota-hungry)
A practical, install-today setup for heavy coding-agent users watching their token bills.
Related event: Developer's Open-Source CLI Cuts Codex Token Usage by 34%(2 posts)→
More from coding & agent
- OpenAI Codex now natively runs open models like GLM-5.3 Flash and Kimi K3 — HarveenChadha · 2026-09-30
- Yacine: Local realtors all know him now — his AI agents retry forever — yacineMTB · 2026-09-30
- Google ships Jetpack Compose A2UI renderer to turn agent streams into native Android UI — rseroter · 2026-09-30
- v0 now lets ChatGPT Plus and Pro subscribers build with their ChatGPT tokens — tomjohndesign · 2026-09-30
- Teaching agents from outcomes: an async reflection loop that turns resolutions into rules — FirstClothes6582 · 2026-09-30
- Vorflux partners with OpenAI: ChatGPT subscriptions now OAuth into its agent harness — ycombinator · 2026-09-30