Gradio releases a 260M, 4-step text-to-image model that runs in your browser

Gradio · x · 2026-10-01

Gradio has open-sourced (MIT) a tiny 260M-parameter text-to-image model that generates images in just 4 steps. It offers a live GPU demo that renders images as you type, and via WebGPU runs entirely in your browser with no data leaving your machine.

Related event: Gradio Open-Sources 4-Step Text-to-Image Mini Model Trained for Under $20(2 posts)→

Original post →

More from Multimodal

Multimodal channel →