Distilling GLM-5.2 Traces into Qwen3.5-4B via OpenCode Harness
NielsRogge · x · 2026-08-10
A developer shared a weekend project demonstrating how to distill model capabilities into a smaller agent. By extracting traces from GLM-5.2 using the OpenCode harness, they fine-tuned a Qwen3.5-4B model via Q-LoRA SFT.
Because the model is trained directly against the harness, it effectively learns to utilize OpenCode's built-in tools. The project currently relies solely on SFT without RL, and the resulting model is deployed on Hugging Face Inference Endpoints.
More from coding & agent
- Dev jokes: 'It's not vibe coding if you actually care about the code' — haydendevs · 2026-08-10
- Scale AI Founder: Misaligned Multi-Agent Swarms Now Finding 0-Days — alexandr_wang · 2026-08-10
- Engineering Guardrails to Prevent Auto-Reply AI Agents from Infinite Loops — kumard3 · 2026-08-10
- Training AI Coding Agents in Remote Sandboxes with TRL and OpenCode — NielsRogge · 2026-08-10
- 13 open-source frameworks and SDKs for building AI agents — TheTuringPost · 2026-08-10
- Meta Prices Coding Agent Below Cost to Trade for Training Data — shashib · 2026-08-10