Qwen 35B Fine-Tune Shows Improvements

habachilles · reddit · 2026-07-13

The author is working on a Qwen 35B fine-tune project, aiming to train on as much "coherent" fable data as possible. Current progress includes: - Renting several H200s for training; - Being highly selective with the data to ensure coherence; - Observing around a +5% improvement on both human eval and SWE benchmarks so far; - Planning to open-source the model and publish metrics after a few more runs. The author also asks the community to contribute fable data, believing these datasets are "already sitting on everyone's hard drives" and can help make local LLMs stronger.

Original post →

More from Models

Models channel →