Tencent Hunyuan's RSR Boosts 27B Model Terminal-Bench 2 pass@3 from 57% to 74%

Tencent-Hunyuan · hf · 2026-10-05

Tencent Hunyuan researchers proposed Recursive Self-Rewrite (RSR), a framework where one base model (Qwen-3.8-27B) discovers successful solutions under diverse harnesses and reconstructs them as training trajectories under a general harness.

Method: a planner extracts procedures into runbooks, a critic screens for verifier/solution leakage and guides recursive revision, and an executor follows qualified runbooks in fresh sandboxes.

Results:

Original post →

More from coding & agent

coding & agent channel →