Dev clones Jev by LoRA-tuning Qwen3.5 4B on 25M synthetic tokens for $2 GPU hours
nato_nob · reddit · 2026-09-21
A developer weekend-cloned a Jev-like model by LoRA fine-tuning Qwen3.5 4B on public datasets plus 25M synthetic tokens generated with DeepSeek V4.1 Flash, training 2 hours on a rented RTX 3090. It lifts typed-decisions from 0.596 to 0.709 — not Jev-level, but clearly stronger than the base model. Everything is open-sourced: weights, synthetic dataset, and a Jev-compatible API endpoint (GitHub: n4ze3m/hmm; HF: n4ze3m/Qwen3.5-4B-Hmm).
More from coding & agent
- Asking my agent to get my sister's Netflix code via her agent, so I never talk to family again — SuB8u · 2026-09-21
- Dev tests model routing: multi-model 'team' costs as much as Astra alone and times out — kevinkern · 2026-09-21
- Should your tool-calling agent get a guardian agent? Devs debate placement and latency — BlueberryOk8225 · 2026-09-21
- Self-hosted Perplexity MCP bridges Perplexity search into ChatGPT on Cloudflare free tier — Sensitive-Priority59 · 2026-09-21
- Open-source Jev sorted 1,000 Gmail messages in 76 seconds for $0.03 — _AustinCalvert_ · 2026-09-21
- Rent the frontier, own the floor: what to hold before the door closes — ccerrato147 · 2026-09-21