Enhancing Human-Object Interaction Editing with I2V Models

机器之心 · wechat · 2026-07-10

A new paper introduces HOI-Edit, a hierarchical cognitive evaluation benchmark designed to assess foundational interaction, spatial understanding, and causal physical reasoning in complex human-object interaction image editing. The authors also propose the HOI-Eval automated evaluation protocol and the SCPE self-correction framework based on I2V models. The related dataset and code are open-source, and the paper has been accepted by ICML 2026.

Original post →

More from Multimodal

Multimodal channel →