$100 full-finetune proves MageFlow 4B's potential as an illustration base model
Turbulent-Bass-649 · reddit · 2026-09-07
A GPU-less university student releases MageTrail, a proof-of-concept full-finetune of Microsoft's MageFlow 4B T2I model on a diversity-maximized 41k-image Danbooru/E621 dataset (curated by Chroma creator Lodestone), tuning booru tag prompting and illustration skills in without the $20k-50k+ cost of a full booru tune.
Key points:
- Dataset recaptioned with Gemini 3.7 Flash (backed by Grok 4.6 and Qwen 3.8 27B FP8), updated to 2026 tag standards; everything open-sourced under a Banodoco grant
- V0.1 trained only 30 epochs (300k samples) for $100; learns booru concepts fast but isn't usable yet (lots of seed rerolling for intact limbs)
- Author argues MageFlow 4B's architecture merits investment: 15-20% faster inference than NVIDIA Cosmos2/Anima despite 2B more params, MageVAE beating QwenVAE, native 256-2048px, MIT license, broad knowledge from 10B curated images, minimal forgetting
- Notes Microsoft seemingly purged the model soon after release, and Krea 2 overshadowed it
- Aiming to raise $700 to reach 200 epochs (V0.5) and attract better-funded investment into the arch
More from Multimodal
- Early testers say Astra's spatial awareness fixes LLMs' chronic scale problem — LinusEkenstam · 2026-09-07
- ShallowStream cuts streaming video understanding latency with shallow-layer index then deep answering — Jitai Hao · 2026-09-07
- GPT-6 Astra builds a 3D site exploding male anatomy into 2,234 modeled pieces — LinusEkenstam · 2026-09-07
- Seedance 2.5 prompt: rapid-fire 0.5s cuts for a found-footage French Polynesia travel video — techhalla · 2026-09-07
- Topaz for Web adds Astra 2 video upscaling, no app or GPU required — azed_ai · 2026-09-07
- How this creator makes anime inserts with Seedance 2.0's reference-video workflow — ring_hyacinth · 2026-09-07