SenseTime releases open-source SenseNova U1.5 with MoT architecture
multimodalart · x · 2026-08-21
SenseTime released the open-source image generation model SenseNova-U1.5-8B-MoT under the Apache 2.0 license. It uses a Mixture of Transformers (MoT) architecture, eliminating the need for VAEs, text encoders, or DiTs. Text and image tokens use different weights and attend to each other, denoising in pixel space. Benchmarks show it matches Nano Banana 2, and a demo is now available.
More from Multimodal
- High-quality icon animations generated via MiniMax H3 — aziz4ai · 2026-08-21
- MiniMax H3 Generates Movie Trailer; 3080Ti Runtime Details Revealed — Working-Distance-901 · 2026-08-21
- AI Video Shifts to Structured Production; Dreamina 2.5 Supports Workflows — CodeByPoonam · 2026-08-21
- Seedance 2.5 shifts AI video from 5-second clips to full film workflows — CodeByPoonam · 2026-08-21
- MiniMax H3 Generates Photorealistic Glacier Collapse Forming 'EXTINCTION' — LudovicCreator · 2026-08-21
- Reddit user highlights issue with DepthAnything V2 output — witcherknight · 2026-08-21