SenseTime’s U1 Pro aims at delivery-grade Chinese visual design

KevinNaughtonJr · x · 2026-07-20

What APPSO tested

APPSO reviews SenseTime’s SenseNova U1 Pro, framing it as more than a text-to-image model: it aims to unify understanding, generation, and action for production-grade visual work.

Key claims from the article

Constraints and caveats

The article also notes that generation can still take 10–30 minutes for complex prompts, and interactive editing is not yet fully open in the first release, though SenseTime says it is being built into the model’s creative loop.

Related event: SenseTime Launches Multimodal Agent Base U1 Pro and Open-Source Vision Dataset(6 posts)→

Original post →

More from Multimodal

Multimodal channel →