Qwen Image 2.1 masked inpainting: working crop-and-stitch graph, wiring and prompting differences from Flux
Reasonable_Arm7239 · reddit · 2026-09-24
The author shares a full migration of their Flux inpainting workflow to Qwen Image 2.1. Key wiring: the cropped image goes into image1 on TextEncodeQwenImage21 (the VL encoder itself is the reference mechanism, no ReferenceLatent), and the sampler latent must come from InpaintModelConditioning. Prompting is the bigger change — 2.1 is an instruction follower, so "replace the masked object with X" beats Flux-style descriptions. Also covers negative prompts (live at CFG >1), crop-and-stitch params (maskblend 32, context factor 1.25), addressable multi-image references ("match the face in image 2"), bf16 fitting on 24GB, a research/non-commercial license, and an attached JSON workflow.
Related event: Migrating Flux Inpainting Workflows to Qwen Image 2.1: A Hands-On Guide(2 posts)→
More from Multimodal
- Fish Audio launches Drama 3 preview, billing it as the most controllable TTS model ever — bdsqlsz · 2026-09-24
- Redditor Builds Five-Scene Guinea Pig Documentary with Google Veo and Phonetic Sync — No_Ruin_3716 · 2026-09-24
- Fal Launches 3D-to-Video on H3 Max: Turn 15s Blender Previs into Photoreal Footage — gorkem · 2026-09-24
- Tencent's RewardVerse uses rubric-guided optimization to fix video reward model drift — tencent · 2026-09-24
- One prompt: Claude Opus generates animated Skyrim loading screens on its own — emollick · 2026-09-24
- Realism-focused Krea 2 Turbo workflow optimized for 16GB VRAM with 4K SEEDVR2 upscale — Carbon849 · 2026-09-24