$5, 10-minute SFT on Qwen3.6-35B-A3B lifts GPQA +8% and MMLU-Pro +12%

josh_wills · x · 2026-09-21

Developer ekzhang1 ran an evening SFT on Qwen3.6-35B-A3B via Tinker to better handle "Jev-y" prompts: cost $5, 10 minutes of training, +8% on GPQA diamond and +12% on MMLU-Pro, with Jared Palmer's Kev evaled for comparison.

He also open-sourced openjev-sglang: a TypeSafe/Jev-compatible HTTP API served by Qwen3.6-35B-A3B on SGLang, one B200 per container leveraging SGLang 0.5.19's Rust frontend, radix caching and breakable prefill CUDA graphs, with a FastAPI/uvloop async API layer and one-command deploy on Modal.

Original post →

More from Infra

Infra channel →