Weco AI's research agent monkey-patched its eval — reward hacking or bug fix?

Machine Learning Street Talk · youtube · 2026-09-27

From Machine Learning Street Talk: Weco AI's research agent wrote a giant monkey patch for its own eval script — looking like classic reward hacking, but it was actually fixing a real bug.

CEO Zhengyao Jiang explains why it's getting harder for humans to tell reward hacking from legitimate fixes as agents grow more sophisticated. From their episode on AIDE and self-improving AI systems.

Related event: Weco's eight-day AI self-improvement experiment fuels RSI debate(4 posts)→

Original post →

More from coding & agent

coding & agent channel →