Weco preprint: agent self-improves its own software 7 times over 8 days, gains transfer to 4 benchmarks
VraserX · x · 2026-09-25
VraserX cites Weco's new preprint to push back on the lazy "it's just a wrapper" framing: the study let an agent autonomously make seven improvements to its own software during an eight-day run, and the gains transferred to four held-out benchmarks — with the underlying models held fixed. The takeaway: improving how an agent searches, manages context and checks its own work can matter a lot on its own, independent of model upgrades.
More from coding & agent
- Building agents for cloud infra: why one bad terraform destroy beats a bad email — Helpful-Man64 · 2026-09-25
- Everyone solved AI code review, nobody solved what happens after merge — trvklhn666 · 2026-09-25
- When AI agents check out off-store, what proof do merchants need to trust it? — ConvertMyStore · 2026-09-25
- Stop Fixing AI Mistakes: Feed Models Your Context, Says Dev Strategy Thread — chaseleantj · 2026-09-25
- fable-advisor ships setup command to max out Claude and ChatGPT subs at once — daniel_mac8 · 2026-09-25
- Lemmalog: a Rust Datalog engine replaces vector stores for LLM agent memory — JeremyCMorgan · 2026-09-25