Dev Complains AI Agents Consume Insane Tokens, Reading 100k Context for Simple Tasks
MiskaMyasa · reddit · 2026-08-04
A developer reports that major LLMs (GPT, Claude Opus, DeepSeek) are exhibiting massive token consumption inflation during coding and agentic tasks.
Simple tasks that previously required minimal context now see models endlessly reading and tracing until the context hits 70K-100K before writing any actual code. Compared to half a year ago, this disconnected resource usage has significantly increased operational costs.
More from coding & agent
- AI Agent Skill Automates 27 Open-Source Docs Generation with Real Context — zyphraxns · 2026-08-04
- Time Management in the Agent Era: Maintaining Human-Machine Asymmetry — TheZachMueller · 2026-08-04
- browser-use releases video-use: editing videos via coding agents — browser-use · 2026-08-04
- Compound Engineering plugin supports Claude Code and Cursor — EveryInc · 2026-08-04
- Uber open-sources ADR: enterprise AI agent security and threat detection framework — uber · 2026-08-04
- Apple's SHARP Model Open-Sourced: Transform Photos into Interactive 3D Scenes — tom_doerr · 2026-08-04