Iris-3B: a 3B pixel-space text-to-image model with no VAE, fully open under Apache 2.0

zhenjun_zhao · x · 2026-10-09

Sperid Labs released Iris-3B, a pixel-space text-to-image model and general vision learner that generates every pixel directly with no VAE.

Related event: Sperid Labs Open-Sources Iris-3B, Challenging Pixel-Space Diffusion(2 posts)→

Original post →

More from Multimodal

Multimodal channel →