Running H3 across a 3090 and unlocked 64GB CMP 170HX hits OOM in ComfyUI
JustinPooDough · reddit · 2026-09-06
A Reddit user combining a 3090 with an unlocked 64GB CMP 170HX to run Minimax H3 (NVFP4-pruned) and Qwen3-VL-32B in ComfyUI keeps hitting OOM on the 3090 with --highvram, despite planning to pin the diffusion model + VAE (46GB) on the CMP and the text encoder (16GB) on the 3090. Without the flag generation takes 15+ minutes due to offloading; the CMP's slow PCI-E 2.0 4x adds constraints. The post asks how to permanently pin components per GPU.
More from Infra
- Blacklisted Inspur Kept Buying Nvidia's Best AI Chips via US Subsidiary Aivres, NYT Finds — ShakeelHashim · 2026-09-06
- Open-source Termux scripts turn old Android phones into GPU-accelerated Linux desktops or Home Assistant hubs — tom_doerr · 2026-09-06
- Grandma GPUs reborn: 2x Tesla P40 hits 48 tok/s on Qwen 27B via F16 cache + MTP — Jumpy-Operation-4615 · 2026-09-06
- Baseten's Philip Kiely Launches Inference Engineering Book, Plus Learning Resources — kmeanskaran · 2026-09-06
- Polygres turns your Postgres into a hybrid search context layer for AI agents — Scobleizer · 2026-09-06
- TCS may invest up to $7.4 billion with TPG in a gigawatt AI campus in Hyderabad — emmanuelvivier · 2026-09-06