Adding 3D Body Scans to LoRA Training
tekprodfx16 · reddit · 2026-07-14
The author forked AI Toolkit to add a second training stage based on SAM 3D Body, enabling LoRA/LoKr to learn not just faces, but accurate body shapes and proportions.
The workflow involves:
- Running standard LoRA/LoKr training to learn trigger words and reference photos
- Generating images and using SAM 3D to scan the bodies in the outputs
- Constraining the training with "real 3D body readings" instead of relying purely on pixels and captions
The author reports a significant improvement in full-body likeness, working seamlessly on Krea 2. Face training takes about 30 minutes, while adding the SAM 3D body stage extends it to roughly 60 minutes. The repo is FedorAiToolkit, and the SAM 3D Body model requires gated access on Hugging Face.
Related event: Enhancing LoRA Training with SAM 3D for Accurate Body Shapes(2 posts)→
More from Multimodal
- Non-coder builds full-featured Android ComfyUI client with ChatGPT, submits to Google Play — ComfierUI · 2026-09-11
- FastH3-Live hits 22fps: acceleration node benchmarks and the --vram-headroom trick — spartong945 · 2026-09-11
- Midjourney style code share: --sref 2912175708 — tisch_eins · 2026-09-11
- Astra storyboards plus Minimax H3 per-shot generation boost video success rates — Hailuo_AI · 2026-09-11
- MiniMax H3 MAX nails cooking anime clips: 15-second curry demo with prompts shared — Hailuo_AI · 2026-09-11
- MiniMax Music Production Toolkit 2.5 for ComfyUI adds full mastering chain — Vivid_Promise1700 · 2026-09-11