Iris-3B: open-weights 3B pixel-space diffusion T2I model with no VAE, Apache 2.0

_akhaliq · x · 2026-10-09

Sperid Labs released Iris-3B, a 3B-parameter text-to-image model trained from scratch entirely in pixel space — generating every pixel directly with no VAE. The paper also explores its generative prior for detail-critical tasks like depth estimation and image restoration. Weights, code, and demo are fully open under Apache 2.0 on Hugging Face.

Related event: Sperid Labs Open-Sources Pixel-Space Text-to-Image Model Iris-3B(4 posts)→

Original post →

More from Multimodal

Multimodal channel →