Dev clones Jev by LoRA-tuning Qwen3.5 4B on 25M synthetic tokens for $2 GPU hours

nato_nob · reddit · 2026-09-21

A developer weekend-cloned a Jev-like model by LoRA fine-tuning Qwen3.5 4B on public datasets plus 25M synthetic tokens generated with DeepSeek V4.1 Flash, training 2 hours on a rented RTX 3090. It lifts typed-decisions from 0.596 to 0.709 — not Jev-level, but clearly stronger than the base model. Everything is open-sourced: weights, synthetic dataset, and a Jev-compatible API endpoint (GitHub: n4ze3m/hmm; HF: n4ze3m/Qwen3.5-4B-Hmm).

Original post →

More from coding & agent

coding & agent channel →