Delving Globally into Texture and Structure for Image Inpainting
Haipeng Liu, Yang Wang, Meng Wang, Yong Rui
摘要
Image inpainting has achieved remarkable progress and inspired abundant methods, where the critical bottleneck is identified as how to fulfill the high-frequency structure and low-frequency texture information on the masked regions with semantics. To this end, deep models exhibit powerful superiority to capture them, yet constrained on the local spatial regions. In this paper, we delve globally into texture and structure information to well capture the semantics for image inpainting. As opposed to the existing arts trapped on the independent local patches, the texture information of each patch is reconstructed from all other patches across the whole image, to match the coarsely filled information, especially the structure information over the masked regions. Unlike the current decoder-only transformer within the pixel level for image inpainting, our model adopts the transformer pipeline paired with both encoder and decoder. On one hand, the encoder captures the texture semantic correlations of all patches across image via self-attention module. On the other hand, an adaptive patch vocabulary is dynamically established in the decoder for the filled patches over the masked regions. Building on this, a structure-texture matching attention module anchored on the known regions comes up to marry the best of these two worlds for progressive inpainting via a probabilistic diffusion process. Our model is orthogonal to the fashionable arts, such as Convolutional Neural Networks (CNNs), Attention and Transformer model, from the perspective of texture and structure information for image inpainting. The extensive experiments over the benchmarks validate its superiority. Our code is available here
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper3
- Towards Coherent Image Inpainting Using Denoising Diffusion Implicit ModelsGuanhua Zhang, Jiabao Ji, Yang Zhang, Mo Yu 等ICML 2023 · 被引用 99 次
- One Stone with Two Birds: A Null-Text-Null Frequency-Aware Diffusion Models for Text-Guided Image InpaintingHaipeng Liu, Yang Wang, Meng WangNeurIPS 2025 · 被引用 8 次
- Structure Matters: Tackling the Semantic Discrepancy in Diffusion Models for Image InpaintingHaipeng Liu, Yang Wang, Biao Qian, Meng Wang 等CVPR 2024
它引用的顶会 Paper16
- An Image is Worth 16x16 Words: Transformers for Image Recognition at ScaleAlexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn 等ICLR 2021 · 被引用 21,477 次
- Free-Form Image Inpainting With Gated ConvolutionJiahui Yu, Zhe Lin, Jimei Yang, Xiaohui Shen 等ICCV 2019 · 被引用 1,990 次
- Generative Pretraining From PixelsMark Chen, Alec Radford, Rewon Child, Jeffrey Wu 等ICML 2020 · 被引用 1,773 次
- TransGAN: Two Pure Transformers Can Make One Strong GAN, and That Can Scale UpYifan Jiang, Shiyu Chang, Zhangyang WangNeurIPS 2021 · 被引用 515 次
- StructureFlow: Image Inpainting via Structure-Aware Appearance FlowYurui Ren, Xiaoming Yu, Ruonan Zhang, Thomas H. Li 等ICCV 2019 · 被引用 356 次
相关 Paper
- Incremental Transformer Structure Enhanced Image Inpainting with Masking Positional EncodingQiaole Dong, Chenjie Cao, Yanwei FuCVPR 2022 · 被引用 194 次
- Atrous Pyramid Transformer with Spectral Convolution for Image InpaintingMuqi Huang, Lefei ZhangACM MM 2022 · 被引用 11 次
- Parallel Multi-Resolution Fusion Network for Image InpaintingWentao Wang, Jianfu Zhang, Li Niu, Haoyu Ling 等ICCV 2021 · 被引用 28 次
- SyFormer: Structure-Guided Synergism Transformer for Large-Portion Image InpaintingJie Wu, Yuchao Feng, Honghui Xu, Chuanmeng Zhu 等AAAI 2024 · 被引用 14 次
- Learning Contextual Transformer Network for Image InpaintingYe Deng, Siqi Hui, Sanping Zhou, Deyu Meng 等ACM MM 2021 · 被引用 29 次
