From-Scratch Model Trained with $1000 Reaches 11% on SWE-bench

sanmikoyejo · x · 2026-08-14

A new study from the Max Planck Institute and Stanford explores the limits of 'speedrunning' SWE-bench on a shoestring budget.

The authors caution that climbing the benchmark without gains in underlying knowledge or math capabilities measures something narrower than general capability.

Related event: Researchers Train SWE-bench Model for $60, Rivaling Claude 2(2 posts)→

Original post →

More from coding & agent

coding & agent channel →