AMD RDNA3 guide: running MiniMax H3 video gen in ComfyUI at ~4 min per 5s clip
Big_Extension_9987 · reddit · 2026-09-10
A Reddit user shared a working optimization guide for running MiniMax H3 video generation locally on an AMD RX 7900 XT (20GB VRAM) with ComfyUI Desktop on Windows 11, after finding AI assistants' advice useless and experimenting manually. Speed improved significantly over defaults with no quality loss.
Performance: 25s/it at 0.8mp (1216x672, 16:9), 4 minutes total for a 5-second text-to-video generation.
Key setup:
- Launch flags: --disable-smart-memory --disable-pinned-memory --disable-triton-backend --use-sage-attention --enable-dynamic-vram
- A set of MIOpen/Triton env vars (logging off, FINDMODE=2, AMD Flash Attention enabled, etc.)
- Quantized 6-step turbo model (14GB w4a8 or 21GB int8 from Hugging Face)
- PlagueKind nodes with SLA Attention; an optional node for longer 15s videos
More from Infra
- 100M output tokens for $60: DeepSeek off-peak pricing undercuts Opus 5 by 40x — airesearch12 · 2026-09-10
- Dev take: token demand will grow far faster than demand for top-line intelligence — willcb · 2026-09-10
- Arm lands Lenovo and ByteDance's Volcengine as first China customers for its AI server chips — pstAsiatech · 2026-09-10
- d-Matrix adopts NVIDIA NVLink Fusion to bring Raptor XPUs to rack-scale deployment — nordicinst · 2026-09-10
- Why DeepSeek might profit despite open weights: it's the only one happy to optimize for its own architecture — yacineMTB · 2026-09-10
- 34-day agent run processed 21.5B input tokens for $200 with 98% cache hits — nodo48 · 2026-09-10