MiniMax H3 Local Generation Fail: Severely Deformed Faces in Videos

Silver-Spot-2763 · reddit · 2026-08-06

A developer reported severe issues on Reddit when using the MiniMax H3 model for text-to-video generation. Although the model is extremely fast, using the int8 quantized pruned version in ComfyUI resulted in horrific facial deformities (like black holes for eyes and frayed edges), reminiscent of early SD 1.5 flaws.

The author noted that Wan and LTX models never exhibited such problems under the same workflow and is currently seeking community help to resolve this generation disaster.

Original post →

More from Multimodal

Multimodal channel →