Learning Global-aware Kernel for Image Harmonization
Xintian Shen, Jiangning Zhang, Jun Chen, Shipeng Bai, Yue Han, Yabiao Wang, Chengjie Wang, Yong Liu
摘要
Image harmonization aims to solve the visual inconsistency problem in composited images by adaptively adjusting the foreground pixels with the background as references. Existing methods employ local color transformation or region matching between foreground and background, which neglects powerful proximity prior and independently distinguishes fore-/back-ground as a whole part for harmonization. As a result, they still show a limited performance across varied foreground objects and scenes. To address this issue, we propose a novel Global-aware Kernel Network (GKNet) to harmonize local regions with comprehensive consideration of long-distance background references. Specifically, GKNet includes two parts, i.e., harmony kernel prediction and harmony kernel modulation branches. The former includes a Long-distance Reference Extractor (LRE) to obtain long-distance context and Kernel Prediction Blocks (KPB) to predict multi-level harmony kernels by fusing global information with local features. To achieve this goal, a novel Selective Correlation Fusion (SCF) module is proposed to better select relevant long-distance background references for local harmonization. The latter employs the predicted kernels to harmonize foreground regions with local and global awareness. Abundant experiments demonstrate the superiority of our method for image harmonization over state-of-the-art methods, e.g., achieving 39.53dB PSNR that surpasses the best counterpart by +0.78dB ↑; decreasing fMSE/MSE by 11.5%↓/6.7%↓ compared with the SoTA method. Code will be available at here. * Equal contribution. † Corresponding author. (a) Local-translation (b) Region-matching (c) Ours
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper2
- High-Resolution Image Harmonization with Adaptive-Interval Color TransformationQuanling Meng, Qinglin Liu, Zonglin Li, Xiangyuan Lan 等NeurIPS 2024 · 被引用 16 次
- Shadow Generation Using Diffusion Model with Geometry PriorHaonan Zhao, Qingyang Liu, Xinhao Tao, Li Niu 等CVPR 2025
它引用的顶会 Paper17
- Swin Transformer: Hierarchical Vision Transformer using Shifted WindowsZe Liu, Yutong Lin, Yue Cao, Han Hu 等ICCV 2021 · 被引用 31,683 次
- An Image is Worth 16x16 Words: Transformers for Image Recognition at ScaleAlexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn 等ICLR 2021 · 被引用 21,477 次
- Generative Pretraining From PixelsMark Chen, Alec Radford, Rewon Child, Jeffrey Wu 等ICML 2020 · 被引用 1,773 次
- High-Fidelity Pluralistic Image Completion with TransformersZiyu Wan, Jingbo Zhang, Dongdong Chen, Jing LiaoICCV 2021 · 被引用 296 次
- FuseFormer: Fusing Fine-Grained Information in Transformers for Video InpaintingRui Liu, Hanming Deng, Yangyi Huang, Xiaoyu Shi 等ICCV 2021 · 被引用 165 次
相关 Paper
- FRIH: Fine-Grained Region-Aware Image HarmonizationJinlong Peng, Zekun Luo, Liang Liu, Boshen ZhangAAAI 2024 · 被引用 18 次
- Image Harmonization with TransformerZonghui Guo, Dongsheng Guo, Haiyong Zheng, Zhaorui Gu 等ICCV 2021 · 被引用 95 次
- Hierarchical Dynamic Image HarmonizationHaoxing Chen, Zhangxuan Gu, Yaohui Li, Jun Lan 等ACM MM 2023 · 被引用 24 次
- DoveNet: Deep Image Harmonization via Domain VerificationWenyan Cong, Jianfu Zhang, Li Niu, Liu Liu 等CVPR 2020
- Deep Image Harmonization with Globally Guided Feature Transformation and Relation DistillationLi Niu, Linfeng Tan, Xinhao Tao, Junyan Cao 等ICCV 2023 · 被引用 14 次
