Lilian Weng's New Blog: Harness Engineering and Recursive Self-Improvement

机器之心 · wechat · 2026-07-07

OpenAI's Lilian Weng published a new blog post, "Harness Engineering for Self-Improvement," returning to a high-frequency posting mode less than 10 days after her last update. The article systematically outlines the "Harness layer" situated between foundation models and the real world—the system architecture responsible for orchestrating model thinking, invoking tools, managing context, and evaluating results. It covers design patterns like workflow automation, file-system persistent memory, and sub-agents, alongside frontier research directions such as ACE, MetaContext Engineering, and the Darwin Godel Machine. Weng explores whether Recursive Self-Improvement (RSI) will first occur at the model weight level or within the Harness layer, candidly listing current bottlenecks: ambiguous evaluation metrics, context and memory lifecycle management, the systematic neglect of negative results, diversity collapse, and reward hacking.

Related event: Lilian Weng and Sakana AI Explore Harness Engineering for RSI(4 posts)→

Original post →

More from coding & agent

coding & agent channel →