Open-World Amodal Appearance Completion
Jiayang Ao, Yanbei Jiang, Qiuhong Ke, Krista A. Ehinger
摘要
Understanding and reconstructing occluded objects is a challenging problem, especially in open-world scenarios where categories and contexts are diverse and unpredictable. Traditional methods, however, are typically restricted to closed sets of object categories, limiting their use in complex, open-world scenes. We introduce Open-World Amodal Appearance Completion, a training-free framework that expands amodal completion capabilities by accepting flexible text queries as input. Our approach generalizes to arbitrary objects specified by both direct terms and abstract queries. We term this capability reasoning amodal completion, where the system reconstructs the full appearance of the queried object based on the provided image and language query. Our framework unifies segmentation, occlusion analysis, and inpainting to handle complex occlusions and generates completed objects as RGBA elements, enabling seamless integration into applications such as 3D reconstruction and image editing. Extensive evaluations demonstrate the effectiveness of our approach in generalizing to novel objects and occlusions, establishing a new benchmark for amodal completion in open-world settings. Code and datasets available: https://github.com/saraao/amodal.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper8
- Amodal3R: Amodal 3D Reconstruction from Occluded 2D ImagesTianhao Wu, Chuanxia Zheng, Frank Guan, Andrea Vedaldi 等ICCV 2025 · 被引用 9 次
- CAPTURe: Evaluating Spatial Reasoning in Vision Language Models via Occluded Object CountingAtin Pothiraj, Elias Stengel-Eskin, Jaemin Cho, Mohit BansalICCV 2025 · 被引用 4 次
- I2E: From Image Pixels to Actionable Interactive Environments for Text-Guided Image EditingJinghan Yu, Junhao Xiao, Chenyu Zhu, Jiaming Li 等ACL 2026 · 被引用 3 次
- SynergyAmodal: Deocclude Anything with Text ControlXinyang Li, Chengjie Yi, Jiawei Lai, Mingbao Lin 等ACM MM 2025 · 被引用 3 次
- Multi-Agent Amodal Completion: Direct Synthesis with Fine-Grained Semantic GuidanceHongxing Fan, Lipeng Wang, Haohua Chen, Zehuan Huang 等ACM MM 2025 · 被引用 3 次
它引用的顶会 Paper17
- Learning Transferable Visual Models From Natural Language SupervisionAlec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh 等ICML 2021 · 被引用 47,906 次
- High-Resolution Image Synthesis with Latent Diffusion ModelsRobin Rombach, Andreas Blattmann, Dominik Lorenz, Patrick Esser 等CVPR 2022 · 被引用 13,123 次
- CogVLM: Visual Expert for Pretrained Language ModelsWeihan Wang, Qingsong Lv, Wenmeng Yu, Wenyi Hong 等NeurIPS 2024 · 被引用 858 次
- Visualizing the Invisible: Occluded Vehicle Segmentation and RecoveryXiaosheng Yan, Yuanlong Yu, Feigege Wang, Wenxi Liu 等ICCV 2019 · 被引用 46 次
- Transparent Image Layer Diffusion using Latent TransparencyLvmin Zhang, Maneesh AgrawalaSIGGRAPH 2024 · 被引用 42 次
相关 Paper
- Variational Amodal Object CompletionHuan Ling, David Acuna, Karsten Kreis, Seung Wook Kim 等NeurIPS 2020 · 被引用 56 次
- Amodal Completion via Progressive Mixed Context DiffusionKatherine Xu, Lingzhi Zhang, Jianbo ShiCVPR 2024 · 被引用 20 次
- Unveiling the Invisible: Reasoning Complex Occlusions Amodally with AURAZhixuan Li, Hyunse Yoon, Sanghoon Lee, Weisi LinICCV 2025
- From Pixels to Logic: A Perception-Reasoning Decomposition Framework for Open-World Referring Expression ComprehensionLihong Huang, Sheng-hua Zhong, Zhi Zhang, Yan LiuAAAI 2026
- Using Diffusion Priors for Video Amodal SegmentationKaihua Chen, Deva Ramanan, Tarasha KhuranaCVPR 2025
