Iris-3B open-sourced: a 3B pixel-space generation and general vision learner
Total-Resort-3120 · reddit · 2026-10-09
A new open-source release, Iris-3B, targets pixel-space generation and general vision learning, working directly in pixel space rather than latent space.
- Weights available on Hugging Face (speridlabs/iris-3b)
- Code and details on GitHub (speridlabs/iris-3b)
- 3B parameters, positioning itself as both a generator and a general vision learner
Related event: Sperid Labs Open-Sources Iris-3B, a Pixel-Space Text-to-Image Model(5 posts)→
More from Multimodal
- 20 GitHub repos turn Claude into a motion design studio — Roger_M_Taylor · 2026-10-09
- Tencent Hunyuan's training-free MC-Sparse attention speeds DiT denoising up to 2.3x — Tencent-Hunyuan · 2026-10-09
- OmniCapBench: 786-video benchmark exposes weak long-horizon audio-visual reasoning — Tencent-Hunyuan · 2026-10-09
- Setting camera paths freely inside generated scenes is a surprisingly cool experience — XRarchitect · 2026-10-09
- Seedance 2.5 AI video demo stuns users with crazy detail level — SimplyAnnisa · 2026-10-09
- Claude Motion ships with day-one HyperFrames Studio integration for video editing — sean_t_strong · 2026-10-09