Musubi Tuner Fork Enables MiniMax H3 LoRA Training on 24GB VRAM

diogodiogogod · reddit · 2026-08-04

A developer shares their fork of Musubi Tuner that enables MiniMax H3 LoRA training on 24GB VRAM (e.g., RTX 4090). By using a pruned ConvRot INT8 model (21GB), BF16 LoRA training is possible with 19-20GB VRAM usage. Tested with 1024x1024 images, batch size 1, LoRA rank/alpha 16, 15 transformer blocks offloaded to CPU, completing two epochs. Advanced features are planned.

Original post →

More from coding & agent

coding & agent channel →