MMFL: Multimodal Fusion Learning for Text-Guided Image Inpainting
Qing Lin, Bo Yan, Jichun Li, Weimin Tan
Abstract
Painters can successfully recover severely damaged objects, yet current inpainting algorithms still can not achieve this ability. Generally, painters will have a conjecture about the seriously missing image before restoring it, which can be expressed in a text description. This paper imitates the process of painters' conjecture, and proposes to introduce the text description into the image inpainting task for the first time, which provides abundant guidance information for image restoration through the fusion of multimodal features. We propose a multimodal fusion learning method for image inpainting (MMFL). To make better use of text features, we construct an image-adaptive word demand module to reasonably filter the effective text features. We introduce a text guided attention loss and a text-image matching loss to make the network pay more attention to the entities in the text description. Extensive experiments prove that our method can better predict the semantics of objects in the missing regions and generate fine grained textures.
Ask about this paper
Ask your agent about it.
Lune has read the top-tier papers around this one, so every answer names the papers it rests on.
Your agent calls
Lunesearch_papers
Free to start. No credit card required.
Terminal
Install the CLIlune papers get 070532a1-4f93-4e21-8d5b-5b3e5d01d309Cited by top-tier papers2
- Text as Neural Operator: Image Manipulation by Text InstructionTianhao Zhang, Hung-Yu Tseng, Lu Jiang, Weilong Yang et al.ACM MM 2021 · 28 citations
- DAFT-GAN: Dual Affine Transformation Generative Adversarial Network for Text-Guided Image InpaintingJihoon Lee, Yunhong Min, Hwidong Kim, Sangtae AhnACM MM 2024 · 3 citations
Related papers
- Text-Guided Image InpaintingZijian Zhang, Zhou Zhao, Zhu Zhang, Baoxing Huai et al.ACM MM 2020 · 17 citations
- Text-Guided Neural Image InpaintingLisai Zhang, Qingcai Chen, Baotian Hu, Shuoran JiangACM MM 2020 · 53 citations
- Adversarial Learning with Mask Reconstruction for Text-Guided Image InpaintingXingcai Wu, Yucheng Xie, Jiaqi Zeng, Zhenguo Yang et al.ACM MM 2021 · 5 citations
- SmartBrush: Text and Shape Guided Object Inpainting with Diffusion ModelShaoan Xie, Zhifei Zhang, Zhe Lin, Tobias Hinz et al.CVPR 2023
- Deep Multi-Resolution Mutual Learning for Image InpaintingHuan Zheng, Zhao Zhang, Haijun Zhang, Yi Yang et al.ACM MM 2022 · 14 citations
