MiniMax H3 Model Launches with Day 0 Native ComfyUI Integration

DavidmComfort · x · 2026-08-04

MiniMax has released the new H3 model, featuring native integration with ComfyUI from day zero.

The model leads with robust multimodal context understanding, capable of processing images, audio, and video simultaneously while resolving them against the prompt. Based on this, H3 collapses multiple video generation tasks into a single model, supporting five core workflows: text-to-video, image-to-video, first-and-last-frame control, reference-to-video, and in-place video editing.

Related event: MiniMax Launches Open-Weight Omni-modal Model H3 with Native 2K Audio-Video Generation(35 posts)→

Original post →

More from Multimodal

Multimodal channel →