SenseNova U1.5 Lite beats FLUX.2-klein-9b on multi-reference image fusion in 5-task test

Kakash1i · reddit · 2026-09-22

A five-task multi-reference image fusion comparison shows SenseNova U1.5 Lite generally making fewer reference errors than FLUX.2-klein-9b-kv, drawing on a wider range of input elements with richer colors and less garbled text.

The biggest gap: in a winter scene prompt requiring the woman's other hand to reach toward the camera catching/tossing a snowball, FLUX.2 kept the original foreground hand holding the ball, while SenseNova rendered the snowball in her outstretched hand, following both references and prompt. All test prompts plus links to both models (black-forest-labs/FLUX.2-klein-9b-kv-fp8, sensenova/SenseNova-U1.5-8B-MoT) are included.

Original post →

More from Multimodal

Multimodal channel →