MiniMax H3 Hybrid Model: Merging First Frame and Reference Images

Tokey_TheBear · reddit · 2026-08-17

A user shared a hybrid model approach for MiniMax H3 (merging FL2VA and REF2VA) to solve the limitation where official models cannot simultaneously maintain high-quality first-frame locking and use extra reference images.

Key Solution:

Use Cases:

Implementation requires specific prompt formatting (e.g., fullypreserved) and ComfyUI wiring using the Reference-to-Video workflow.

Original post →

More from Multimodal

Multimodal channel →