Bespoke Labs post-trains Inkling with SFT+GRPO, gaining 57pp on repo-level coding and 40% token efficiency

AlexGDimakis · x · 2026-09-04

Bespoke Labs researchers post-trained the Inkling base model on a specific GitHub repo using SFT on strong-teacher trajectories plus GRPO RL in repo-specialized environments. SFT alone added 52pp on the held-out fontTools eval; RL lifted it to 57pp, with good transfer to Terminal-Bench 2.1 and SWE-Bench Lite and 40% better token efficiency.

Original post →

More from coding & agent

coding & agent channel →