Continuously Masked Transformer for Image Inpainting
Keunsoo Ko, Chang-Su Kim
摘要
A novel continuous-mask-aware transformer for image inpainting, called CMT, is proposed in this paper, which uses a continuous mask to represent the amounts of errors in tokens. First, we initialize a mask and use it during the self-attention. To facilitate the masked self-attention, we also introduce the notion of overlapping tokens. Second, we update the mask by modeling the error propagation during the masked self-attention. Through several masked self-attention and mask update (MSAU) layers, we predict initial inpainting results. Finally, we refine the initial results to reconstruct a more faithful image. Experimental results on multiple datasets show that the proposed CMT algorithm outperforms existing algorithms significantly. The source codes are available at https://github.com/keunsoo-ko/CMT.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper9
- EditMGT: Unleashing Potentials of Masked Generative Transformers in Image EditingWei Chow, Linfeng Li, Lingdong Kong, Zefeng Li 等CVPR 2026 · 被引用 14 次
- BLS-GAN: A Deep Layer Separation Framework for Eliminating Bone Overlap in Conventional RadiographsHaolin Wang, Yafei Ou, Prasoon Ambalathankandy, Gen Ota 等AAAI 2025 · 被引用 7 次
- Perspective-Aware 3D Gaussian Inpainting with Multi-View ConsistencyYuxin Cheng, Binxiao Huang, Taiqiang Wu, Wenyong Zhou 等ICCV 2025 · 被引用 1 次
- Structure Matters: Tackling the Semantic Discrepancy in Diffusion Models for Image InpaintingHaipeng Liu, Yang Wang, Biao Qian, Meng Wang 等CVPR 2024
- RORem: Training a Robust Object Remover with Human-in-the-LoopRuibin Li, Tao Yang, Song Guo, Lei ZhangCVPR 2025
它引用的顶会 Paper13
- Swin Transformer: Hierarchical Vision Transformer using Shifted WindowsZe Liu, Yutong Lin, Yue Cao, Han Hu 等ICCV 2021 · 被引用 31,683 次
- An Image is Worth 16x16 Words: Transformers for Image Recognition at ScaleAlexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn 等ICLR 2021 · 被引用 21,477 次
- Free-Form Image Inpainting With Gated ConvolutionJiahui Yu, Zhe Lin, Jimei Yang, Xiaohui Shen 等ICCV 2019 · 被引用 1,990 次
- MAT: Mask-Aware Transformer for Large Hole Image InpaintingWenbo Li, Zhe Lin, Kun Zhou, Lu Qi 等CVPR 2022 · 被引用 382 次
- Large Scale Image Completion via Co-Modulated Generative Adversarial NetworksShengyu Zhao, Jonathan Cui, Yilun Sheng, Yue Dong 等ICLR 2021 · 被引用 348 次
相关 Paper
- Learning Contextual Transformer Network for Image InpaintingYe Deng, Siqi Hui, Sanping Zhou, Deyu Meng 等ACM MM 2021 · 被引用 29 次
- Reduce Information Loss in Transformers for Pluralistic Image InpaintingQiankun Liu, Zhentao Tan, Dongdong Chen, Qi Chu 等CVPR 2022 · 被引用 99 次
- DLFormer: Discrete Latent Transformer for Video InpaintingJingjing Ren, Qingqing Zheng, Yuanyuan Zhao, Xuemiao Xu 等CVPR 2022 · 被引用 39 次
- Diverse Image Inpainting with Bidirectional and Autoregressive TransformersYingchen Yu, Fangneng Zhan, Rongliang Wu, Jianxiong Pan 等ACM MM 2021 · 被引用 153 次
- TransCNN-HAE: Transformer-CNN Hybrid AutoEncoder for Blind Image InpaintingHaoru Zhao, Zhaorui Gu, Bing Zheng, Haiyong ZhengACM MM 2022 · 被引用 30 次
