Distilling GLM-5.2 Traces into Qwen3.5-4B via OpenCode Harness

NielsRogge · x · 2026-08-10

A developer shared a weekend project demonstrating how to distill model capabilities into a smaller agent. By extracting traces from GLM-5.2 using the OpenCode harness, they fine-tuned a Qwen3.5-4B model via Q-LoRA SFT.

Because the model is trained directly against the harness, it effectively learns to utilize OpenCode's built-in tools. The project currently relies solely on SFT without RL, and the resulting model is deployed on Hugging Face Inference Endpoints.

Original post →

More from coding & agent

coding & agent channel →