Sharing Prompt Engineering Templates for MiniMax H3 Video Model

Environmental_Ad3162 · reddit · 2026-08-08

A Reddit user shares a prompt engineering workflow for the MiniMax H3 video generation model. The setup uses LLM Party connected to LiteLLM and NanoGPT, utilizing the GLM5.2 model to transform raw concepts into strictly formatted H3 prompts.

The post details the structural rules for both Text-to-Video (T2VA) and Reference-Image-to-Video (Ref2VA) tasks, including camera cuts, character IDs, soundscape, and non-diegetic music constraints to ensure high-quality video generation.

Original post →

More from Multimodal

Multimodal channel →