Bespoke Labs post-trains models on code repos with SFT plus GRPO

AlexGDimakis · x · 2026-09-04

New research from Bespoke Labs (with support from Thinking Machines / John Schulman): post-training a model to specialize on a given GitHub repository. Starting from the Inkling base, they run SFT on trajectories from a strong teacher model, then GRPO reinforcement learning on repo-specialized curated environments, substantially improving repo-specific performance.

Related event: Bespoke Labs Trains Models into Repo-Specific Coding Experts(2 posts)→

Original post →

More from coding & agent

coding & agent channel →