MiniMax H3 local test on RTX 5060ti: 800s inference
erioca · reddit · 2026-08-17
A user test of the MiniMax H3 model (ref2va workflow) running on older hardware (i7 6700 + RTX 5060ti 16GB + 32GB DDR3). Compared the hybrid model with the int8 pruned version, noting the hybrid offers better visuals but often imports white backgrounds from character sheets. Detailed settings include 16:9 resolution at 0.6 ratio and 15-second duration. Average inference time was 800 seconds, utilizing Turbo LoRA and various attention optimization patches.
Related event: Hands-On Tests Show MiniMax H3 Runs Locally on Consumer GPUs(2 posts)→
More from Multimodal
- Indic-Transcribe: Speech model supporting 26 Indian languages launches — sumanthd17 · 2026-08-17
- LTX 2.5 Workflow: Perfect Lip Sync with Custom Audio — False_Suspect_6432 · 2026-08-17
- Creator ships episode 8 of self-made AI thriller built in InVideo and ElevenLabs Music — bennash · 2026-08-17
- Seedance 2.5 Generates Photorealistic Tokyo Travel Vlog with Strong Identity Consistency — eyishazyer · 2026-08-17
- Open-source H3 Prompt Studio: local LLM writes MiniMax H3 storyboards — lololerigolo60 · 2026-08-17
- Connection Error visual combining Grok Imagine and Seedance 2.0 — creatoroff · 2026-08-17