Running H3 across a 3090 and unlocked 64GB CMP 170HX hits OOM in ComfyUI

JustinPooDough · reddit · 2026-09-06

A Reddit user combining a 3090 with an unlocked 64GB CMP 170HX to run Minimax H3 (NVFP4-pruned) and Qwen3-VL-32B in ComfyUI keeps hitting OOM on the 3090 with --highvram, despite planning to pin the diffusion model + VAE (46GB) on the CMP and the text encoder (16GB) on the 3090. Without the flag generation takes 15+ minutes due to offloading; the CMP's slow PCI-E 2.0 4x adds constraints. The post asks how to permanently pin components per GPU.

Original post →

More from Infra

Infra channel →