Learning Contextual Transformer Network for Image Inpainting
Ye Deng, Siqi Hui, Sanping Zhou, Deyu Meng, Jinjun Wang
摘要
Fully Convolutional Networks with attention modules have been proven effective for learning-based image inpainting. While many existing approaches could produce visually reasonable results, the generated images often show blurry textures or distorted structures around corrupted areas. The main reason is due to the fact that convolutional neural networks have limited capacity for modeling contextual information with long range dependencies. Although the attention mechanism can alleviate this problem to some extent, existing attention modules tend to emphasize similarities between the corrupted and the uncorrupted regions while ignoring the dependencies from within each of them. Hence, this paper proposes the Contextual Transformer Network (CTN) which not only learns relationships between the corrupted and the uncorrupted regions but also exploits their respective internal closeness. Besides, instead of a fully convolutional network, in our CTN, we stack several transformer blocks to replace convolution layers to better model the long range dependencies. Finally, by dividing the image into patches of different sizes, we propose a multi-scale multi-head attention module to better model the affinity among various image regions. Experiments on several benchmark datasets demonstrate superior performance by our proposed approach.
问问这篇 Paper
问问你的智能体。
Lune 读过与它相关的顶会 Paper,每个回答都会注明依据哪几篇。
引用它的顶会 Paper4
- T-former: An Efficient Transformer for Image InpaintingYe Deng, Siqi Hui, Sanping Zhou, Deyu Meng 等ACM MM 2022 · 被引用 58 次
- Contextual Outpainting with Object-Level Contrastive LearningJiacheng Li, Chang Chen, Zhiwei XiongCVPR 2022 · 被引用 10 次
- View-consistent Object Removal in Radiance FieldsYiren Lu, Jing Ma, Yu YinACM MM 2024 · 被引用 3 次
- Instruct2See: Learning to Remove Any Obstructions Across DistributionsJunhang Li, Yu Guo, Chuhua Xian, Shengfeng HeICML 2025
相关 Paper
- Atrous Pyramid Transformer with Spectral Convolution for Image InpaintingMuqi Huang, Lefei ZhangACM MM 2022 · 被引用 11 次
- TransCNN-HAE: Transformer-CNN Hybrid AutoEncoder for Blind Image InpaintingHaoru Zhao, Zhaorui Gu, Bing Zheng, Haiyong ZhengACM MM 2022 · 被引用 30 次
- MAT: Mask-Aware Transformer for Large Hole Image InpaintingWenbo Li, Zhe Lin, Kun Zhou, Lu Qi 等CVPR 2022 · 被引用 382 次
- Incremental Transformer Structure Enhanced Image Inpainting with Masking Positional EncodingQiaole Dong, Chenjie Cao, Yanwei FuCVPR 2022 · 被引用 194 次
- Frequency-Aware Spatiotemporal Transformers for Video Inpainting DetectionBingyao Yu, Wanhua Li, Xiu Li, Jiwen Lu 等ICCV 2021 · 被引用 38 次
