Scale AI launches RSI-Bench: $2,000 per accepted task plus co-authorship for contributors

dhruv2038 · x · 2026-08-18

Scale AI has launched RSI-Bench, an ongoing effort to evaluate whether AI agents can develop the capabilities required to advance AI R&D — a prerequisite for recursive self-improvement. Each task ships with a starting environment, a compute budget, and a verifier suite measuring performance against a baseline. Beyond final outcomes, the benchmark evaluates reliability (distinguishing real gains from noise), efficiency (solving under limited resources), generality (solutions that extend beyond narrow fixes), and idea quality (extracting insights and synthesizing novel approaches).

The team is crowdsourcing expert-curated, long-horizon research tasks from the community. Contributors receive co-authorship on the RSI-Bench paper, $2,000 per accepted task for the initial 50 tasks, Modal compute credits, and access to the research community.

Original post →

More from Research

Research channel →