Before Diffusion: Looking Back at VQGAN, the Pre-Diffusion Foundation of Image Generation
makeitrad1 · x · 2026-09-11
@makeitrad1 shares a "Before Diffusion" piece revisiting VQGAN — the technique that encoded images as discrete tokens and laid groundwork for generative vision before diffusion models took over.
More from Multimodal
- Dancers, rain, and a mirror-floor dive: video shot entirely by an AI model — tsi_org · 2026-09-11
- Amap's ABot-Earth 0.7 Generates a Roamable 3D City in 10 Minutes from One Satellite Image — 量子位 · 2026-09-11
- Turn any article into a podcast with Meta's Muse — alexandr_wang · 2026-09-11
- GPT-6 Astra 3D Workflow: Blender MCP for Hard-Surface, TripoAI for Organic Models — majidmanzarpour · 2026-09-11
- ComfyUI node brings a 3-light 3D dome relighting studio to MiniMax H3 Edit — Emotional_Example_12 · 2026-09-11
- Optimized video workflow: 10s at 1MP in ~125s on an RTX 5090, with custom audio driving — Tokyo_Jab · 2026-09-11