Google's RRSI regularizes recursive self-improvement of agent harnesses

Google's RRSI paper adds regularization to recursive self-improvement of agent harnesses, gaining 4.7 points on out-of-distribution tasks while saving about 30% of tokens.

2026-09-22 ~ 2026-09-23 · 4 related posts