Fudan and CUHK MMLab publish first survey of Agentic Visual Generation, classifying systems L0-L4

jiqizhixin · x · 2026-10-01

Fudan University's Wan Team and CUHK MMLab have released the first survey of Agentic Visual Generation, addressing the shift from "prompt in, image out" toward systems that autonomously decide what to generate, pick tools mid-execution, revise from results, and retain experience.

Original post →

More from Multimodal

Multimodal channel →