Weekend project turns small Qwen3.5 fine-tunes into top decision-index models
Individual-Stay-4193 · reddit · 2026-10-06
Developer jiwidi shared Jiwo, a weekend project turning small LLMs into decision models, releasing two Qwen3.5 fine-tunes (0.8B and 4B).
- On the Decision Index leaderboard shared last week, the 0.8B tops the sub-1B category, 45.7% above the best other Qwen3.5-0.8B fine-tune; the 4B ranks second in the 3–6B class
- Most of the work was data calibration: sourcing external datasets and re-adapting them to decision scenarios; the author expects the community to push numbers higher, especially if Qwen 3.8 releases arrive for these sizes
- Code is open-sourced on GitHub; the leaderboard runs on a Hugging Face Space
More from Models
- OpenAI to watermark ChatGPT text outputs to comply with EU AI Act — OpenAI · 2026-10-06
- ChatGPT Adds Another Disclaimer About Self-Serving Rankings in B2B — lilyraynyc · 2026-10-06
- SelfBench turns real GitHub PRs into evals: open-weight models cost more and do worse — ycombinator · 2026-10-06
- Embedded Grok on X reportedly lacks per-user context isolation, called out as a major flaw — altryne · 2026-10-06
- Liquid AI's d1 decision model adds vision, beats GPT-6.1 on 4 of 6 tasks at up to 200x lower cost — JosephJacks_ · 2026-10-06
- Anthropic reviewers alerted police to a Claude chat threatening a sheriff's office, leading to an arrest — rohanpaul_ai · 2026-10-06