Codex Repeatedly Compresses and Burns Tokens in Long Tasks

aigclink · x · 2026-07-19

A post pointed out that Codex can insanely burn through tokens in long tasks due to a flaw in its compression design: after compression, the model only sees the initial instructions, causing it to repeatedly enter an "initial instructions -> compress -> execute initial instructions" loop. It might only complete the task by chance if it skips compression; otherwise, it drains the entire quota. The accompanying image further illustrates this runaway loop: once the context is compressed, the working memory is repeatedly reset, leading to redundant execution and uncontrolled token usage. The poster tagged relevant personnel, hoping they would look into the issue.

Related event: Alibaba Announces 2.4T Open-Weight Model Qwen3.8(21 posts)→

Original post →

More from coding & agent

coding & agent channel →