OmniColor: A Unified Framework for Multi-modal Lineart Colorization (ECCV 2026)
机器之心 · wechat · 2026-08-27
Accepted by ECCV 2026, the paper OmniColor proposes a unified multi-modal lineart colorization framework to handle the alignment and conflict of mixed control signals (e.g., text, color hints, reference images, history frames) in animation production. The core innovation categorizes control signals into two types: spatial alignment conditions (lineart, color hints, recent frames) and semantic reference conditions (text, identity reference, long-term frames). The model uses a dual-encoder strategy for spatial conditions, introduces a TRE module to eliminate redundancy in history frames, and an AS-Gate module to dynamically handle conflicts. Experiments show that the method outperforms existing approaches in quality, control precision, and sequence consistency, supporting flexible control combinations in real production workflows.
More from Multimodal
- Meitu MT Lab Presents CFT for Stable Portrait Relighting at ECCV 2026 — jiqizhixin · 2026-08-27
- Flux.3 Tested with 8 Reference Images: Quality Degradation? — R34vspec · 2026-08-27
- 15-Second 768p Video Generated in 5.12 Seconds — umesh_ai · 2026-08-27
- Visual General Intelligence White Paper: Vision as a Pathway to AGI — zhenjun_zhao · 2026-08-27
- Video generation speeds: 23.7s vs 11m shows Jevons Paradox in action — gorkem · 2026-08-27
- V-Rubrics: Visual Faithfulness via Rubric-Based RL — nanyang-technological-university-singapore · 2026-08-27