Fix anatomy in text-to-video with H3 'anatomical slider' prompt technique

SIR_NVAX_A_LOT · reddit · 2026-08-29

The author shares a technique for maintaining anatomical consistency in text-to-video (t2v) generation using a so-called "anatomical slider" within H3 prompts. By adding detailed descriptions of the reference frame—covering lighting, composition, subject features, and clothing—the model can keep the subject stable throughout the video. A detailed prompt example specifying cinematic style, lighting direction, and character attire is provided.

Original post →

More from Multimodal

Multimodal channel →