SenseTime Releases U1Pro Multimodal Creation Model

新智元 · wechat · 2026-07-18

At WAIC, SenseTime released its next-generation multimodal model, SenseNova U1Pro, featuring a unified base for "understanding, generation, and action." The article primarily showcases its image generation capabilities in tasks like 8K scrolls, movie posters, infographics, Chinese long-text typography, and character design sheets, emphasizing that it acts more like a "thinking designer" rather than a slot-machine-style image generator.

Technically, U1Pro employs a proprietary unified architecture called NEO-unify, integrating understanding, generation, and action into the same representation and Transformer sequence. The creation process is completed through long-range iterations similar to an Agentic Generation Loop. The article also notes:

Related event: SenseTime Launches SenseNova U1 Pro Native 8K Multimodal Model(6 posts)→

Original post →

More from Models

Models channel →