Qwen-Image-3.0 gets put through layout-heavy tests against GPT-Image-2

Scobleizer · x · 2026-07-23

Qwen-Image-3.0 is being benchmarked against GPT-Image-2 on dense-layout tasks

A repost from Scobleizer compares Qwen-Image-3.0 with GPT-Image-2 on image-generation tasks that stress “useful” outputs rather than just pretty ones.

The test suite is based on Alibaba Qwen’s own claims for the model: rich multi-element layouts, legible text, and practical use cases such as newspaper PDFs, storyboards, exam papers, and product layouts.

Tasks tested

Early results mentioned in the thread

The quoted Alibaba Qwen post says Qwen-Image-3.0 is a third-generation foundational image model built around one goal: “Real” — richer content, authentic details, and deeper knowledge, including support for 4.5k-token prompts, 10px-readable text, 12 languages, and realistic UI rendering.

Original post →

More from Models

Models channel →