H3 Character LoRA Training: Likeness Far Behind Wan 2.2, Tips and Pitfalls Shared
Tiny-Highlight-9180 · reddit · 2026-09-09
A Reddit user shares experience training an H3 character LoRA with disappointing results. Using an RTX 5090, musubi-tuner, 60 images, 3000 steps (5 hours). Key hyperparameters: networkdim/alpha=16, LR 1e-6, sigmoid timestep sampling, EMA decay 0.999, blockstoswap=38. Testing 1500 vs 3000 steps shows 1500 is insufficient; EMA vs non-EMA made little difference. Other findings: fl2va and ref2va are separate, LoRAs don't transfer; enabling usepinnedmemoryforblockswap cut iteration time from 11.6s to 6.1s. The author notes the same dataset produced much better results on Wan 2.2 and seeks advice.
More from Multimodal
- Grok Imagine Adds First and Last Frame Control for Video Generation — mark_k · 2026-09-09
- AI-Generated Video of a Cat Riding a Roller Coaster at Santa Monica Pier — Born_History_8246 · 2026-09-09
- Tencent Hunyuan open-sources AuK, a foundational model unifying speech generation and editing — Tencent-Hunyuan · 2026-09-09
- Video gen is so fast that 'script' may no longer be the right pre-production artifact — tobowers · 2026-09-09
- ComfyUI Style Node NeonsStyleExplorer Adds Custom Catalogs and Style Libraries — neonsparksuk · 2026-09-09
- DBZ x High School of the Dead crossover AI animation, artstyle drift and all — Ok-Giraffe-8670 · 2026-09-09