Adding 3D Body Scans to LoRA Training
tekprodfx16 · reddit · 2026-07-14
The author forked **AI Toolkit** to add a second training stage based on **SAM 3D Body**, enabling LoRA/LoKr to learn not just faces, but accurate body shapes and proportions. The workflow involves: 1. Running standard LoRA/LoKr training to learn trigger words and reference photos 2. Generating images and using SAM 3D to scan the bodies in the outputs 3. Constraining the training with "real 3D body readings" instead of relying purely on pixels and captions The author reports a significant improvement in **full-body likeness**, working seamlessly on Krea 2. Face training takes about 30 minutes, while adding the SAM 3D body stage extends it to roughly 60 minutes. The repo is `FedorAiToolkit`, and the SAM 3D Body model requires gated access on Hugging Face.
Related event: Enhancing LoRA Training with SAM 3D for Accurate Body Shapes(2 posts)→
More from Multimodal
- AI-made “cat mode” stunt turns a skateboard clip into a surreal landing demo — taherdhanera · 2026-07-21
- OCT-Bench sets 10,076 questions to test whether multimodal models really understand retinal scans — Baochen Fu · 2026-07-21
- LTX-2.3 face-and-voice LoRA training can work on 12GB VRAM with heavy tradeoffs — __alpha_____ · 2026-07-21
- Seedance 2.0 turns one reference image into a cinematic fight scene — techhalla · 2026-07-21
- Seedance 2.0 keeps character consistency across 15+ shots with just 3 prompts — techhalla · 2026-07-21
- DecartAI’s Lucy 2.5 Realtime lands on fal with live video-to-video editing — gorkem · 2026-07-21