Burn Bar for Omarchy visualizes Claude/Codex token burn, quotas and GPU load locally
DanWahlin · x · 2026-09-05
A new Burn Bar for Omarchy shows every model your machine is burning in real time: token usage bucketed in 30-minute windows, split into input, cache write, output and cache reads; per-plan quota bars with countdowns and reset times that pulse above 90%; per-model Claude spend; plus GPU load, watts, temperature, VRAM share and which model is resident, with buttons to warm or unload models.
The implementation is deliberately minimal: a Python 3 collector with zero third-party dependencies, no API key and no telemetry — it just reads the transcripts Claude Code and Codex already write to disk. The only network call is to your own Ollama.
More from coding & agent
- Dev says GPT-6 Astra solved his months-stuck mocap retargeting workflow in two hours — tinyfool · 2026-09-05
- Open-source AI trading agent turns social sentiment into signals via Gemini — tom_doerr · 2026-09-05
- A dev building an Epub plugin breaks down the evolution of TTF, OTF and WOFF2 — vista8 · 2026-09-05
- Codex community plans ~40 global meetups with hackathons and multi-agent workshops — gabrielchua · 2026-09-05
- Magnitude: open-source server picks the best local models for your hardware and plugs into your coding agent — solyarisoftware · 2026-09-05
- Dev patches WebKitGTK CVE to restore drag-and-drop, ships Copilot app as Flatpak — unixterminal · 2026-09-05