MiniMax H3 Video Model Repurposed as Image Generator with Surprising Design Skills

sktksm · reddit · 2026-08-05

Developer sktksm shared a ComfyUI workflow that repurposes the MiniMax H3 video model as a pseudo-image generator.

Instead of native text-to-image, H3 generates a short sequence, decodes it into an image batch via the video VAE, and extracts a single frame as the final still image.

Test Results:

Recommended Parameters: The author found that an INT Length of 8 and an Image From Batch Index of 8 yielded the best results, though users are encouraged to preview the full batch and test nearby values based on the prompt.

Related event: MiniMax H3 Video Model Adapted for Image Generation(2 posts)→

Original post →

More from Multimodal

Multimodal channel →