ROME's training pipeline: 500B-token CPT, error-masked SFT, chunk-level RL

thisguyknowsai · x · 2026-10-06

A breakdown of ROME's three-stage training pipeline:

Each stage builds on the last — systematic capability building with no shortcuts.

Related event: Chinese Team Open-Sources ROME+ALE Agent Ecosystem; 30B Sparse Model Claims Parity with 480B+ Rivals(9 posts)→

Original post →

More from coding & agent

coding & agent channel →