Adding 3D Body Scans to LoRA Training

tekprodfx16 · reddit · 2026-07-14

The author forked **AI Toolkit** to add a second training stage based on **SAM 3D Body**, enabling LoRA/LoKr to learn not just faces, but accurate body shapes and proportions. The workflow involves: 1. Running standard LoRA/LoKr training to learn trigger words and reference photos 2. Generating images and using SAM 3D to scan the bodies in the outputs 3. Constraining the training with "real 3D body readings" instead of relying purely on pixels and captions The author reports a significant improvement in **full-body likeness**, working seamlessly on Krea 2. Face training takes about 30 minutes, while adding the SAM 3D body stage extends it to roughly 60 minutes. The repo is `FedorAiToolkit`, and the SAM 3D Body model requires gated access on Hugging Face.

Related event: Enhancing LoRA Training with SAM 3D for Accurate Body Shapes(2 posts)→

Original post →

More from Multimodal

Multimodal channel →