Custom ComfyUI Node Boosts Qwen-Image 2.1 Prompt Enhancer 1.4-1.7x via MTP Head
mozophe · reddit · 2026-09-25
A developer found ComfyUI's official Qwen-Image 2.1 prompt enhancer (a Qwen3.5-9B fine-tune) slow due to a missing MTP head. Attaching an MTP head boosted throughput on an RTX 4090 Laptop 16GB: 35→47 tok/s for t2i (1.36x) and 31→52 tok/s for editing (1.65x); vs the official maxlength checkpoint, up to 1.7x faster for t2i and 2.5x for editing, with unchanged quality (same seed yields different text).
Since ComfyUI needed MTP-specific fixes for Qwen 3.5 image input, the author released a custom node: ComfyUI-Qwen-Image-2.1-PromptEnhancer-MTP.
Details:
- Requires ComfyUI v0.37.0+, 16GB VRAM recommended (peak 14GB t2i, 16GB editing), 32GB RAM, 10GB disk per model
- Auto-downloads and attaches the MTP head on first use; uses Qwen's official system prompts and sampling settings
- Includes heretic (uncensored) versions with the same speed-up
- Supports editing with up to 10 input images; prompt and reasoning output separately; sample workflows included
More from Infra
- VAST Data: the $30B hidden software layer of NVIDIA's AI stack, CEO interview — mattturck · 2026-09-25
- RTX 5090's 32GB VRAM vs 256k Context: Is It Enough for Local Qwen 27B Agentic Coding? — lots_of_puppies · 2026-09-25
- yetone teases 'One more thing' LLM gateway, dev installs it on day one — vista8 · 2026-09-25
- Lightmatter CEO: moving lasers onto 300mm silicon wafers to unlock optical interconnect — BenBajarin · 2026-09-25
- Oracle Says Datacenter Force-Majeure Notice Doesn't Mean Project Delay — ns123abc · 2026-09-25
- 100MW+ Datacenters Wait 5-10 Years for Grid Power; Cato Says Let Firms Build Their Own — McDonaghMatthew · 2026-09-25