Supra2-IMG: 100M-Param Text-to-Image Model Trained in Under 10 Hours on One H100, Open-Sourced
LH-Tech_AI · reddit · 2026-09-21
SupraLabs released Supra2-IMG, an open-source 100M-parameter DiT text-to-image model trained from scratch in under 10 hours on a single Runpod H100, claiming SOTA quality at 256×256.
- All sample images use identical settings: seed 0, 50 steps, cfg 3.0 — no cherry-picking, per the authors
- Extremely cheap inference: 20s per image on CPU, 2s on GPU
- A one-line wget of inference.py gets you local generation running
- Model available on Hugging Face (SupraLabs/Supra2-IMG)
The standout is pushing text-to-image down to a size that runs on plain CPUs — interesting for tiny-model research and on-device generation.
More from Multimodal
- Astra is a breakthrough for VFX work where other AI video tools fall short — danshipper · 2026-09-22
- Workflow: generate a video, then edit and animate it freely in Blender's Grease Pencil — andrew_n_carr · 2026-09-22
- User finds Qwen 2.1 doubles as a near-SeedVR2-level upscaler in image workflow — skyrimer3d · 2026-09-22
- Grok prompt turns your avatar into a stitched circular patch design — bennash · 2026-09-22
- Dev builds AI Scene Director on JEV: one prompt rewrites Three.js scene — jamestagg · 2026-09-22
- Musician Jordi Pons releases track made with AI sound design, not AI songwriting — jordiponsdotme · 2026-09-22