Huawei & CUHK open-source Lego-RL: plug real coding agent harnesses into RL, SWE-bench 64.0→70.4

青稞AI · wechat · 2026-09-06

Huawei and CUHK researchers open-sourced Lego-RL, a Harness-Native RL training framework for coding agents that plugs real harnesses (OpenHands SDK, Claude Code, OpenCode) into RL without changing a single line of their code. On Qwen3.5-35B-A3B, SWE-bench Verified scores rose from 64.0/62.4/57.2 to 70.4/68.2/66.6 across three harnesses.

The framework tackles three failure modes of RL on real agent harnesses:

Paper: arxiv.org/abs/2608.17393; code: github.com/LegoX/Lego-RL. Core developer Du Yiming (Huawei Leibniz Institute, CUHK PhD) will present the work in a Qingke AI talk on Sept 8.

Original post →

More from coding & agent

coding & agent channel →