Keeping character scale consistent in MiniMax H3 Ref2V: prompting tricks that mostly fail
fluvialcrunchy · reddit · 2026-09-16
A creator working with MiniMax H3 Ref2V reference-to-video finds relative scale between characters/objects hard to keep consistent — even within the same shot.
Prompting like '<subject 1> is slightly taller than <subject 2>' has limited success. They ask whether the model can read a scale bar or centimeter measurements from a character sheet, and note first/last frame references help but don't fit every project. Seeking community best practices.
More from Multimodal
- GMI launches MCP server exposing 150+ multimodal models to Claude, ChatGPT and Cursor — _jaydeepkarale · 2026-09-16
- Audio8 open-sources on-device ASR/TTS models down to 0.1B, including iPhone offline transcription — FinanceYF5 · 2026-09-16
- GPT Image 2.5 versus seven other image models on the same prompt — ZootAllures9111 · 2026-09-16
- FP8+AOTI-optimized Wan2.2 video model space trends on Hugging Face — zerogpu-aoti · 2026-09-16
- MiniMax Unveils H3 IP Edition with Officially Licensed Japanese IP — MiniMax_AI · 2026-09-16
- Cartwheel MCP connects AI animation to Unreal Engine — andrew_n_carr · 2026-09-16