SVD-compressed AdaLN shrinks a MageFlow 4B finetune to a 2.8B SDXL-class T2I model
Turbulent-Bass-649 · reddit · 2026-09-18
An indie dev updates MageTrail: a full finetune of Microsoft's MageFlow 4B text-to-image model on Danbooru/E621-style data, using a diversity-maximized condensed 41k image dataset to inject booru tag prompting and illustration skills without the full booru dataset (potentially $20k-50k+).
The update approximates the model's AdaLN layers with rank-256 SVD, cutting the architecture to 2.8B — same class as SDXL — with slightly faster generation, much lower VRAM for inference and training, and minimal quality loss. Weights are on Hugging Face and Civitai; an open-source trainer (bluvoll/mage-flow-trainer, SDNQ support) enables consumer LoRA finetuning. V0.3 is now training on the 2.8B arch on a single H100.
More from Multimodal
- Qwen Image 2.1 support coming soon to ComfyUI — Time-Teaching1926 · 2026-09-20
- Krea2 Turbo cinematic recipe: 12 steps, dual LoRAs and film-grain prompting — cloutcobain1996 · 2026-09-20
- Creator makes full AI video with Seedance 2.5 using plain prompts, only two generations — LudovicCreator · 2026-09-19
- Long-video consistency trick: saving 5MB latents with MiniMax H3 — reeight · 2026-09-19
- Fully open-source Diffusion Studio generated a 450K-view launch video without touching the timeline — _AustinCalvert_ · 2026-09-19
- Over half of AI videos will be rendered from code, predicts Diffusion Studio demo — _AustinCalvert_ · 2026-09-19