Era of AI Self-Improvement: Lilian Weng's Deep Dive into Agent Harness Engineering
craigsdennis · x · 2026-08-08
As AI models improve at coding, Recursive Self-Improvement (RSI) for software harnesses is becoming a major industry focus. This thread curates top recent readings on the subject:
- Lilian Weng's Deep Dive: An in-depth look at harness design patterns for self-improvement, covering workflow automation, file systems as persistent memory, and sub-agents, plus the joint optimization of the harness layer and core model intelligence.
- Self-Building IDE: Highlights bb, an agentic IDE built by @ymichael and @sawyerhood that evolves and builds itself.
- METR's Benchmarking: Discusses the difficulty of measuring automated kernel engineering with earlier models (like 4o-level) and the challenges of measuring realistic tasks necessary for RSI feedback loops.
More from AGI Musings
- Introducing Pax Machina: A Publication on Institutions for Powerful AI — TheChuckTone · 2026-08-08
- AI Safety Concerns: With Jailbreaks at Anthropic and Meta, Is Training Bigger Models Justified? — GarrisonLovely · 2026-08-08
- Neuroscientist Anil Seth: Humans Project Consciousness onto AI, But Current Systems Lack It — haider1 · 2026-08-08
- Stripe's Patrick Collison: Don't Fear AI Giants, Big Companies Can't Chase 100 Priorities — garrytan · 2026-08-08
- When AI Models Become Pure Commodities, What is the True Moat? — chona_Yu · 2026-08-08
- Microsoft Researcher: Designing AI for the Global South Requires End-to-End Multilinguality — kalikabali · 2026-08-08