TransCNN-HAE: Transformer-CNN Hybrid AutoEncoder for Blind Image Inpainting
Haoru Zhao, Zhaorui Gu, Bing Zheng, Haiyong Zheng
Abstract
Blind image inpainting is extremely challenging due to the unknown and multi-property complexity of contamination in different contaminated images. Current mainstream work decomposes blind image inpainting into two stages: mask estimating from the contaminated image and image inpainting based on the estimated mask, and this two-stage solution involves two CNN-based encoder-decoder architectures for estimating and inpainting separately. In this work, we propose a novel one-stage Transformer-CNN Hybrid AutoEncoder (TransCNN-HAE) for blind image inpainting, which intuitively follows the inpainting-then-reconstructing pipeline by leveraging global long-range contextual modeling of Transformer to repair contaminated regions and local short-range contextual modeling of CNN to reconstruct the repaired image. Moreover, a Cross-layer Dissimilarity Prompt (CDP) is devised to accelerate the identifying and inpainting of contaminated regions. Ablation studies validate the efficacy of both TransCNN-HAE and CDP, and extensive experiments on various datasets with multi-property contaminations show that our method achieves state-of-the-art performance with much lower computational cost on blind image inpainting. Our code is available at https://github.com/zhenglab/TransCNN-HAE.
Ask about this paper
Ask your agent about it.
Lune has read the top-tier papers around this one, so every answer names the papers it rests on.
Your agent calls
Lunesearch_papers
Free to start. No credit card required.
Terminal
Install the CLIlune papers get 09fd8dd1-1fde-4d9d-8da9-b109f7efa577Cited by top-tier papers3
- BVINet: Unlocking Blind Video Inpainting With Zero AnnotationsZhiliang Wu, Kerui Chen, Kun Li, Hehe Fan et al.ICCV 2025 · 30 citations
- Text Image Inpainting via Global Structure-Guided Diffusion ModelsShipeng Zhu, Pengfei Fang, Chenjie Zhu, Zuoyan Zhao et al.AAAI 2024 · 25 citations
- Reproducing the Past: A Dataset for Benchmarking Inscription RestorationShipeng Zhu, Hui Xue, Na Nie, Chenjie Zhu et al.ACM MM 2024 · 4 citations
Related papers
- Learning Contextual Transformer Network for Image InpaintingYe Deng, Siqi Hui, Sanping Zhou, Deyu Meng et al.ACM MM 2021 · 29 citations
- Image Harmonization with TransformerZonghui Guo, Dongsheng Guo, Haiyong Zheng, Zhaorui Gu et al.ICCV 2021 · 95 citations
- Diverse Image Inpainting with Bidirectional and Autoregressive TransformersYingchen Yu, Fangneng Zhan, Rongliang Wu, Jianxiong Pan et al.ACM MM 2021 · 153 citations
- Context-Aware Pretraining for Efficient Blind Image DecompositionChao Wang, Zhedong Zheng, Ruijie Quan, Yifan Sun et al.CVPR 2023
- SyFormer: Structure-Guided Synergism Transformer for Large-Portion Image InpaintingJie Wu, Yuchao Feng, Honghui Xu, Chuanmeng Zhu et al.AAAI 2024 · 14 citations
