How to Manage Excessively Long Agent Contexts
Kevin_Dong_cn · reddit · 2026-07-11
The post asks: when managing agent conversation context, how should one handle the massive amount of context data generated by tool calls in each turn? The author currently only retains user questions, tool names, and answers, avoiding dumping the full tool output directly because schema results and queries can be too long. However, including all tool history and outputs would cause rapid bloat, potentially reaching **1 million tokens** by the 5th turn. They are looking to learn how the community typically handles context compression, truncation, or storage management.
Related event: Developers Debate Context Bloat in Data Agents(2 posts)→
More from Infra
- Larry Fink says China is ahead in the AI energy race, citing 100 GW nuclear buildout — rohanpaul_ai · 2026-07-21
- Local AI may pay back in 6–7 years and cut long-term costs by 30–40% — DavidLinthicum · 2026-07-21
- TSMC reportedly plans up to 10% chipmaking price hikes in 2027 — kimmonismus · 2026-07-21
- More open models and llama.cpp updates are coming, says Merve Noyan — mervenoyann · 2026-07-21
- Why adding a second LLM provider breaks more than the API surface — Ok_Extension6373 · 2026-07-21
- UK AI datacentres face backlash over heat, noise and land use — nordicinst · 2026-07-21