Merging video and image LoRAs remains unsolved; most combos are inconsistent

linoy_tsaban · x · 2026-10-05

AI practitioner Linoy Tsaban raises an industry pain point: we still lack a good way to merge LoRAs, for both video and image generation. She acknowledges that merging overlapping or contradicting LoRAs is semantically ambiguous and understandably hard, but notes that aside from fast-inference LoRA + other LoRA combos that usually work well, most other combinations produce highly inconsistent results.

Original post →

More from Multimodal

Multimodal channel →