Precise Object and Effect Removal with Adaptive Target-Aware Attention
Jixin Zhao, Zhouxia Wang, Peiqing Yang, Shangchen Zhou
Abstract
Object removal requires eliminating not only the target object but also its associated visual effects such as shadows and reflections. However, diffusion-based inpainting and removal methods often introduce artifacts, hallucinate contents, alter background, and struggle to remove object effects accurately. To address these challenges, we propose ObjectClear, a novel framework that decouples foreground removal from background reconstruction via an adaptive target-aware attention mechanism. This design empowers the model to precisely localize and remove both objects * Equal contribution † Corresponding author and their effects while maintaining high background fidelity. Moreover, the learned attention maps are leveraged for an attention-guided fusion strategy during inference, further enhancing visual consistency. To facilitate the training and evaluation, we construct OBER, a large-scale dataset for OBject-Effect Removal, which provides paired images with and without object-effects, along with precise masks for both objects and their effects. The dataset comprises high-quality captured and simulated data, covering diverse objects, effects, and complex multi-object scenes. Extensive experiments demonstrate that ObjectClear outperforms prior methods, achieving superior object-effect removal quality and background fidelity, especially in challenging scenarios.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext eaf96161-39ea-4784-916d-b1ce2d78d8d7Cited by top-tier papers6
- EffectErase: Joint Video Object Removal and Insertion for High-Quality Effect ErasingYANG FU, Yike Zheng, Ziyun Dai, Henghui DingCVPR 2026 · 14 citations
- Refaçade: Editing Object with Given Reference TextureYouze Huang, Penghui Ruan, Bojia Zi, Xianbiao Qi et al.CVPR 2026 · 3 citations
- 3D Space as a Scratchpad for Editable Text-to-Image GenerationOindrila Saha, Vojtech Krs, Radomír Mech, Subhransu Maji et al.CVPR 2026 · 1 citation
- In-Context Generation with Regional Constraints for Instructional Video EditingZhongwei Zhang, Fuchen Long, Wei Li, Zhaofan Qiu et al.ICML 2026
- BFS: Back-to-Front Layered Image Synthesis via Knowledge TransferKyoungkook Kang, Gyujin Sim, Sunghyun ChoSIGGRAPH 2026
Builds on28
- Learning Transferable Visual Models From Natural Language SupervisionAlec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh et al.ICML 2021 · 47,906 citations
- Denoising Diffusion Probabilistic ModelsJonathan Ho, Ajay Jain, Pieter AbbeelNeurIPS 2020 · 35,902 citations
- Segment AnythingAlexander Kirillov, Eric Mintun, Nikhila Ravi, Hanzi Mao et al.ICCV 2023 · 13,211 citations
- High-Resolution Image Synthesis with Latent Diffusion ModelsRobin Rombach, Andreas Blattmann, Dominik Lorenz, Patrick Esser et al.CVPR 2022 · 13,123 citations
- SDXL: Improving Latent Diffusion Models for High-Resolution Image SynthesisDustin Podell, Zion English, Kyle Lacey, Andreas Blattmann et al.ICLR 2024 · 4,569 citations
Related papers
- GeoRemover: Removing Objects and Their Causal Visual ArtifactsZixin Zhu, Haoxiang Li, Xuelu Feng, He Wu et al.NeurIPS 2025 · 11 citations
- PS-Diffusion: Photorealistic Subject-Driven Image Editing with Disentangled Control and AttentionWeicheng Wang, Guoli Jia, Zhongqi Zhang, Liang Lin et al.CVPR 2025
- Attentive Eraser: Unleashing Diffusion Model's Object Removal Potential via Self-Attention Redirection GuidanceWenhao Sun, Xue-Mei Dong, Benlei Cui, Jingqun TangAAAI 2025 · 50 citations
- CLIPAway: Harmonizing focused embeddings for removing objects via diffusion modelsYigit Ekin, Ahmet Burak Yildirim, Erdem Eren Caglar, Aykut Erdem et al.NeurIPS 2024 · 26 citations
- ROSE: Remove Objects with Side Effects in VideosChenxuan Miao, Yutong Feng, Jianshu Zeng, Zixiang Gao et al.NeurIPS 2025 · 37 citations
