Testing MiniMax H3: Generating High-Consistency Yu-Gi-Oh Anime with 9 Reference Images

circlenline · reddit · 2026-08-05

Using the MiniMax H3 video generation model (ref2v mode) locally on an RTX 5080, a developer successfully generated a 14-second Yu-Gi-Oh themed animation using only 9 reference images.

The test demonstrated exceptional multi-subject consistency:

The author also shared specific local deployment configs, including FP8 pruned diffusion models, Qwen3VL text encoder, and sampler parameters.

Original post →

More from Multimodal

Multimodal channel →