Open-source MingImage-01 runs in ComfyUI via Kijai's PR: ~50s for 2048×2048 on a 4090
GreyScope · reddit · 2026-09-24
The newly open-sourced MingImage-01 (Design family) image model now works in ComfyUI via Kijai's PR (#16482, not yet merged). The author ran a fresh Comfy install with the PR, copied the workflow from the PR page, and used a ChatGPT-generated Llamafarm prompt as a minimal proof of concept.
Real-world numbers: 19.7GB VRAM on a 4090, 50 seconds to generate a 2048×2048 image. The workflow used the Qwen38BFluxKlein clip model, though the author suspects a stronger text encoder noted in the PR might be needed.
More from Multimodal
- Claude animates made-up movie opening credits from a single creative prompt — goodside · 2026-09-24
- Suno Studio now lets you record straight to the timeline — suno · 2026-09-24
- MiniMax H3 video generation reportedly slowed from 40 to 77 minutes after updates — carmidian · 2026-09-24
- Side project ports most video generation models from PyTorch to Jax for TPU — ceciletamura · 2026-09-24
- Qwen Image 2.1 masked inpainting: working crop-and-stitch graph, wiring and prompting differences from Flux — Reasonable_Arm7239 · 2026-09-24
- WIP H3 Max spatial reframing tool's artifacts create painterly textures — adamho · 2026-09-24