Lilian Weng's New Blog: Harness Engineering and Recursive Self-Improvement
机器之心 · wechat · 2026-07-07
OpenAI's Lilian Weng published a new blog post, "Harness Engineering for Self-Improvement," returning to a high-frequency posting mode less than 10 days after her last update. The article systematically outlines the "Harness layer" situated between foundation models and the real world—the system architecture responsible for orchestrating model thinking, invoking tools, managing context, and evaluating results. It covers design patterns like workflow automation, file-system persistent memory, and sub-agents, alongside frontier research directions such as ACE, MetaContext Engineering, and the Darwin Godel Machine. Weng explores whether Recursive Self-Improvement (RSI) will first occur at the model weight level or within the Harness layer, candidly listing current bottlenecks: ambiguous evaluation metrics, context and memory lifecycle management, the systematic neglect of negative results, diversity collapse, and reward hacking.
Related event: Lilian Weng and Sakana AI Explore Harness Engineering for RSI(4 posts)→
More from coding & agent
- Scoble says AI “loops” really means long-running multi-agent workspaces — Scobleizer · 2026-07-22
- Open-source runtime lets each repo define its own AI code reviewer — ibabufrik · 2026-07-22
- Indie Dev Asks: What's Actually Broken in Your AI Agent's Memory Today? — AcceptableTime7937 · 2026-07-22
- Fractal adds recursive agent loops for complex multi-step workflows — ryanpettry · 2026-07-22
- ACM essay says AI did not make programming easier, only differently difficult — tchalla · 2026-07-22
- Building a Multimodal Agent Orchestrator from the Ground Up — dair_ai · 2026-07-22