MiniMax H3 local test on RTX 5060ti: 800s inference

erioca · reddit · 2026-08-17

A user test of the MiniMax H3 model (ref2va workflow) running on older hardware (i7 6700 + RTX 5060ti 16GB + 32GB DDR3). Compared the hybrid model with the int8 pruned version, noting the hybrid offers better visuals but often imports white backgrounds from character sheets. Detailed settings include 16:9 resolution at 0.6 ratio and 15-second duration. Average inference time was 800 seconds, utilizing Turbo LoRA and various attention optimization patches.

Related event: Hands-On Tests Show MiniMax H3 Runs Locally on Consumer GPUs(2 posts)→

Original post →

More from Multimodal

Multimodal channel →