Amodal3R: Amodal 3D Reconstruction from Occluded 2D Images
Tianhao Wu, Chuanxia Zheng, Frank Guan, Andrea Vedaldi, Tat-Jen Cham
Abstract
Most existing image-to-3D models assume that objects are fully visible, ignoring occlusions that commonly occur in real-world scenarios. In this paper, we introduce Amodal3R, a conditional image-to-3D model designed to reconstruct plausible 3D geometry and appearance from partial observations. We extend a “foundation” 3D generator by introducing a visible mask-weighted attention mechanism and an occlusion-aware attention layer that explicitly leverage visible and occlusion priors to guide the reconstruction process. We demonstrate that, by training solely on synthetic data, Amodal3R learns to recover full 3D objects even in the presence of occlusions in real scenes. It substantially outperforms state-of-the-art methods that independently perform amodal completion followed by reconstruction, thereby establishing a new benchmark for occlusion-aware 3D reconstruction.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers18
- SAM 3D: 3Dfy Anything in ImagesXingyu Chen, Fu-Jen Chu, Pierre Gleize, Kevin J Liang et al.CVPR 2026 · 280 citations
- ShapeR: Robust Conditional 3D Shape Generation from Casual CapturesYawar Siddiqui, Duncan P. Frost, Samir Aroudj, Armen Avetisyan et al.CVPR 2026 · 19 citations
- HoloScene: Simulation-Ready Interactive 3D Worlds from a Single VideoHongchi Xia, Chih-Hao Lin, Hao-Yu Hsu, Quentin Leboutet et al.NeurIPS 2025 · 18 citations
- SceneMaker: Open-set 3D Scene Generation with Decoupled De-occlusion and Pose Estimation ModelYukai Shi, Weiyu Li, Zihao Wang, Hongyang Li et al.CVPR 2026 · 17 citations
- Orientation Matters: Making 3D Generative Models Orientation-AlignedYichong Lu, Yuzhuo Tian, Zijin Jiang, Yikun Zhao et al.NeurIPS 2025 · 15 citations
Builds on56
- Learning Transferable Visual Models From Natural Language SupervisionAlec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh et al.ICML 2021 · 47,906 citations
- Denoising Diffusion Probabilistic ModelsJonathan Ho, Ajay Jain, Pieter AbbeelNeurIPS 2020 · 35,902 citations
- Segment AnythingAlexander Kirillov, Eric Mintun, Nikhila Ravi, Hanzi Mao et al.ICCV 2023 · 13,211 citations
- High-Resolution Image Synthesis with Latent Diffusion ModelsRobin Rombach, Andreas Blattmann, Dominik Lorenz, Patrick Esser et al.CVPR 2022 · 13,123 citations
- 3D Gaussian Splatting for Real-Time Radiance Field RenderingBernhard Kerbl, Georgios Kopanas, Thomas Leimkühler, George DrettakisSIGGRAPH 2023 · 5,687 citations
Related papers
- Using Diffusion Priors for Video Amodal SegmentationKaihua Chen, Deva Ramanan, Tarasha KhuranaCVPR 2025
- RnG: A Unified Transformer for Complete 3D Modeling from Partial ObservationsMochu Xiang, Zhelun Shen, Xuesong li, Jiahui Ren et al.CVPR 2026 · 2 citations
- Variational Amodal Object CompletionHuan Ling, David Acuna, Karsten Kreis, Seung Wook Kim et al.NeurIPS 2020 · 56 citations
- Open-World Amodal Appearance CompletionJiayang Ao, Yanbei Jiang, Qiuhong Ke, Krista A. EhingerCVPR 2025
- Amodal Ground Truth and Completion in the WildGuanqi Zhan, Chuanxia Zheng, Weidi Xie, Andrew ZissermanCVPR 2024 · 23 citations
