SVD-compressed AdaLN shrinks a MageFlow 4B finetune to a 2.8B SDXL-class T2I model

Turbulent-Bass-649 · reddit · 2026-09-18

An indie dev updates MageTrail: a full finetune of Microsoft's MageFlow 4B text-to-image model on Danbooru/E621-style data, using a diversity-maximized condensed 41k image dataset to inject booru tag prompting and illustration skills without the full booru dataset (potentially $20k-50k+).

The update approximates the model's AdaLN layers with rank-256 SVD, cutting the architecture to 2.8B — same class as SDXL — with slightly faster generation, much lower VRAM for inference and training, and minimal quality loss. Weights are on Hugging Face and Civitai; an open-source trainer (bluvoll/mage-flow-trainer, SDNQ support) enables consumer LoRA finetuning. V0.3 is now training on the 2.8B arch on a single H100.

Original post →

More from Multimodal

Multimodal channel →