Qwen Image Edit appears trained at 1MP: resizing inputs to 1024x1024 yields far better results

Civil_Fee_7862 · reddit · 2026-08-26

The author noticed Qwen Image Edit performs substantially better when images are resized to 1024x1024 during encoding, then upscaled back after generation. The speculation is that the model was trained on 1MP inputs, but no docs confirm this. The author also notes 1MP seems optimal for other models as well.

Original post →

More from Multimodal

Multimodal channel →