Agent Now Picks Its Own Effort and Hands Off to Subagents, Leading Cost-Performance on DeepSWE and Terminal-Bench 4.0

_jzhao · x · 2026-09-30

The team behind the Agent describes two updates: the agent now dynamically chooses its own reasoning effort, and its ability to hand off work to subagents has improved. Together these put it, by the authors' claim, at the frontier of the cost-performance curve on DeepSWE and Terminal-Bench 4.0.

The author adds a take straight out of the Bitter Lesson: bet on the model getting smarter, and make sure the harness isn't getting in the way.

Original post →

More from coding & agent

coding & agent channel →