FrogNano: 4B model hits 61.5% on SWE-bench Verified with pure RL, zero distillation

sivareddyg · x · 2026-09-09

FrogNano is a Qwen3.5-4B model post-trained purely with RL on synthetic tasks from TaskPilot — no distillation, no teacher-trajectory SFT. Just 5 iterations × 300 tasks got it to 61.5% on SWE-bench Verified.

Related event: FrogNano: 4B Model Trained with Pure RL Hits 61.5% on SWE-bench(2 posts)→

Original post →

More from coding & agent

coding & agent channel →