Should agents internalize the loop, or keep enforcement outside?
DanielKhashabi · x · 2026-07-21
A thread argues that the next step after reasoning models and learned tool use is to internalize the agent loop itself.
The proposed direction is for models to own planning, tool use, error recovery, verification, memory, and stopping decisions—possibly even tool design and skill composition. But the author draws a hard boundary: the model should handle cognition, while the harness still provides sandboxing, permissions, budgets, and auditing as an enforcement layer that does not trust the policy.
The core question is where that boundary should sit, and how to distinguish true internalization from superficial tool-form behavior.
More from coding & agent
- Anthropic researcher: 99% of engineers now run swarms of 300+ self-improving agents — AlishaOutridge · 2026-09-11
- Gergely Orosz: Shipping 10x PRs With AI Agents, Sites Fill With Small Regressions — ducha_aiki · 2026-09-11
- Same Echo Maze prompt, three frontier models: all passed visually but shipped the same hidden bug — eyishazyer · 2026-09-11
- Astra storyboards plus Minimax H3 per-shot generation boost video success rates — Hailuo_AI · 2026-09-11
- Codex tip: use Sol with Astra and Luna sub-agents to save usage — pvncher · 2026-09-11
- agents-best-practices: a provider-neutral Agent Skill for designing and auditing agentic harnesses — tom_doerr · 2026-09-11