Qwen 2.5 VL fine-tuning: DoRA merge produces base-model-like output in Unsloth

Double-Primary-2871 · reddit · 2026-09-17

A Reddit user fine-tuning Qwen 2.5 VL with bf16 on a single RX 7900 XT (ROCm) reports that while the training loop runs fine, test inference or merging the DoRA weights yields responses identical to the base model, as if no fine-tuning happened.

Their comparison: LoRA works with minimal anchoring at conservative learning rates, but DoRA shows almost no effect even with a modest learning rate, despite usedora=True in the config. They ask whether anyone has successfully merged a DoRA adapter using Unsloth.

Original post →

More from Models

Models channel →