UMM-Reflection: Interleaved RL Teaches Unified Multimodal Models to Self-Correct Images

Ziqi Huang's team open-sourced UMM-Reflection, which uses interleaved reinforcement learning to teach unified multimodal models to generate, reflect on, and redraw images in a joint loop. The method improves GenEval scores from 0.71 to 0.84.

2026-09-29 ~ 2026-09-29 · 3 related posts

1 near-duplicate retellings: ziqi_huang_