SenseNova U1.5 Lite beats FLUX.2-klein-9b on multi-reference image fusion in 5-task test
Kakash1i · reddit · 2026-09-22
A five-task multi-reference image fusion comparison shows SenseNova U1.5 Lite generally making fewer reference errors than FLUX.2-klein-9b-kv, drawing on a wider range of input elements with richer colors and less garbled text.
The biggest gap: in a winter scene prompt requiring the woman's other hand to reach toward the camera catching/tossing a snowball, FLUX.2 kept the original foreground hand holding the ball, while SenseNova rendered the snowball in her outstretched hand, following both references and prompt. All test prompts plus links to both models (black-forest-labs/FLUX.2-klein-9b-kv-fp8, sensenova/SenseNova-U1.5-8B-MoT) are included.
More from Multimodal
- AI-generated bull statue swaps into a museum, fooling viewers online — umesh_ai · 2026-09-22
- AI influencer vlogs level up: 30-second consistent clips with same face and story — aftahi_ai · 2026-09-22
- Canva Gets 70% More Video Generations per GPU Hour on NVIDIA Blackwell — pbaylies · 2026-09-22
- Manually adjust camera angles with Viggle's H3 Meridian, for both image and video — linoy_tsaban · 2026-09-22
- Startup claims a neuron-inspired software layer makes AI video 5x faster and 80% cheaper — The Decoder · 2026-09-22
- AI-Generated Short Film Depicts a Cinematic 'Glass' Delivery Rider — NoVeterinarian5438 · 2026-09-22