ROME hits 57.4% on SWE-bench Verified with only 3B activated parameters

thisguyknowsai · x · 2026-10-06

Benchmark results for the ROME model:

The author claims this 30B model matches 480B+ models on real coding tasks, challenging assumptions about agent scaling.

Related event: Chinese Team Open-Sources ROME+ALE Agent Ecosystem; 30B Sparse Model Claims Parity with 480B+ Rivals(9 posts)→

Original post →

More from coding & agent

coding & agent channel →