Hierarchical Dynamic Image Harmonization
Haoxing Chen, Zhangxuan Gu, Yaohui Li, Jun Lan, Changhua Meng, Weiqiang Wang, Huaxiong Li
摘要
Image harmonization is a critical task in computer vision, which aims to adjust the foreground to make it compatible with the background. Recent works mainly focus on using global transformations (i.e., normalization and color curve rendering) to achieve visual consistency. However, these models ignore local visual consistency and their huge model sizes limit their harmonization ability on edge devices. In this paper, we propose a hierarchical dynamic network (HDNet) to adapt features from local to global view for better feature transformation in efficient image harmonization. Inspired by the success of various dynamic models, local dynamic (LD) module and mask-aware global dynamic (MGD) module are proposed in this paper. Specifically, LD matches local representations between the foreground and background regions based on semantic similarities, then adaptively adjust every foreground local representation according to the appearance of its 𝐾-nearest neighbor background regions. In this way, LD can produce more realistic images at a more fine-grained level, and simultaneously enjoy the characteristic of semantic alignment. The MGD effectively applies distinct convolution to the foreground and background, learning the representations of foreground and background regions as well as their correlations to the global harmonization, facilitating local visual consistency for the images much more efficiently. Experimental results demonstrate that the proposed HDNet significantly reduces the total model parameters by more than 80% compared to previous methods, while still attaining state-of-the-art performance on the popular iHarmony4 dataset. Additionally, we introduced a lightweight version of HDNet, i.e., HDNet-lite, which has only 0.65MB parameters, yet it still achieved competitive performance. Our code is avaliable in https://github.com/chenhaoxing/HDNet.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper6
- Zero-shot Image Editing with Reference ImitationXi Chen, Yutong Feng, Mengting Chen, Yiyang Wang 等NeurIPS 2024 · 被引用 80 次
- DiffUTE: Universal Text Editing Diffusion ModelHaoxing Chen, Zhuoer Xu, Zhangxuan Gu, Jun Lan 等NeurIPS 2023 · 被引用 61 次
- Deep Image Harmonization in Dual Color SpacesLinfeng Tan, Jiangtong Li, Li Niu, Liqing ZhangACM MM 2023 · 被引用 21 次
- Training-and-Prompt-Free General Painterly Harmonization via Zero-Shot Disentenglement on Style and Content ReferencesTeng-Fang Hsiao, Bo-Kai Ruan, Hong-Han ShuaiAAAI 2025 · 被引用 6 次
- AnyDoor: Zero-shot Object-level Image CustomizationXi Chen, Lianghua Huang, Yu Liu, Yujun Shen 等CVPR 2024
它引用的顶会 Paper14
- EditGAN: High-Precision Semantic Image EditingHuan Ling, Karsten Kreis, Daiqing Li, Seung Wook Kim 等NeurIPS 2021 · 被引用 248 次
- Dynamic Instance Normalization for Arbitrary Style TransferYongcheng Jing, Xiao Liu, Yukang Ding, Xinchao Wang 等AAAI 2020 · 被引用 212 次
- Image Harmonization with TransformerZonghui Guo, Dongsheng Guo, Haiyong Zheng, Zhaorui Gu 等ICCV 2021 · 被引用 95 次
- High-Resolution Image Harmonization via Collaborative Dual TransformationsWenyan Cong, Xinhao Tao, Li Niu, Jing Liang 等CVPR 2022 · 被引用 86 次
- SCS-Co: Self-Consistent Style Contrastive Learning for Image HarmonizationYucheng Hang, Bin Xia, Wenming Yang, Qingmin LiaoCVPR 2022 · 被引用 49 次
相关 Paper
- Learning Global-aware Kernel for Image HarmonizationXintian Shen, Jiangning Zhang, Jun Chen, Shipeng Bai 等ICCV 2023 · 被引用 14 次
- FRIH: Fine-Grained Region-Aware Image HarmonizationJinlong Peng, Zekun Luo, Liang Liu, Boshen ZhangAAAI 2024 · 被引用 18 次
- Region-Aware Adaptive Instance Normalization for Image HarmonizationJun Ling, Han Xue, Li Song, Rong Xie 等CVPR 2021
- PCT-Net: Full Resolution Image Harmonization Using Pixel-Wise Color TransformationsJulian Jorge Andrade Guerreiro, Mitsuru Nakazawa, Björn StengerCVPR 2023
- Deep Image Harmonization with Learnable AugmentationLi Niu, Junyan Cao, Wenyan Cong, Liqing ZhangICCV 2023 · 被引用 13 次
