Shape Anchor Guided Holistic Indoor Scene Understanding
Mingyue Dong, Linxi Huan, Hanjiang Xiong, Shuhan Shen, Xianwei Zheng
摘要
This paper proposes a shape anchor guided learning strategy (AncLearn) for robust holistic indoor scene under-standing. We observe that the search space constructed by current methods for proposal feature grouping and instance point sampling often introduces massive noise to instance detection and mesh reconstruction. Accordingly, we develop AncLearn to generate anchors that dynamically fit instance surfaces to (i) unmix noise and target-related features for offering reliable proposals at the detection stage, and (ii) reduce outliers in object point sampling for directly providing well-structured geometry priors without segmentation during reconstruction. We embed AncLearn into a reconstruction-from-detection learning system (AncRec) to generate high-quality semantic scene models in a purely instance-oriented manner. Experiments conducted on the challenging ScanNetv2 dataset demonstrate that our shape anchor-based method consistently achieves state-of-the-art performance in terms of 3D object detection, layout estimation, and shape reconstruction. The code will be available at https://github.com/Geo-Tell/AncRec.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper2
- LASA: Instance Reconstruction from Real Scans using A Large-scale Aligned Shape Annotation DatasetHaolin Liu, Chongjie Ye, Yinyu Nie, Yingfan He 等CVPR 2024 · 被引用 1 次
- PanoContext-Former: Panoramic Total Scene Understanding with a TransformerYuan Dong, Chuan Fang, Liefeng Bo, Zilong Dong 等CVPR 2024
它引用的顶会 Paper17
- End-to-End CAD Model Retrieval and 9DoF Alignment in 3D ScansArmen Avetisyan, Angela Dai, Matthias NießnerICCV 2019 · 被引用 88 次
- VENet: Voting Enhancement Network for 3D Object DetectionQian Xie, Yu-Kun Lai, Jing Wu, Zhoutao Wang 等ICCV 2021 · 被引用 60 次
- Towards Part-Based Understanding of RGB-D ScansAlexey Bokhovkin, Vladislav Ishimtsev, Emil Bogomolov, Denis Zorin 等CVPR 2021
- Total3DUnderstanding: Joint Layout, Object Pose and Mesh Reconstruction for Indoor Scenes From a Single ImageYinyu Nie, Xiaoguang Han, Shihui Guo, Yujian Zheng 等CVPR 2020
- ImVoteNet: Boosting 3D Object Detection in Point Clouds With Image VotesCharles R. Qi, Xinlei Chen, Or Litany, Leonidas J. GuibasCVPR 2020
相关 Paper
- SPGroup3D: Superpoint Grouping Network for Indoor 3D Object DetectionYun Zhu, Le Hui, Yaqi Shen, Jin XieAAAI 2024 · 被引用 24 次
- Learning 3D Scene Priors with 2D SupervisionYinyu Nie, Angela Dai, Xiaoguang Han, Matthias NießnerCVPR 2023
- 3D Instance Segmentation via Multi-Task Metric LearningJean Lahoud, Bernard Ghanem, Martin R. Oswald, Marc PollefeysICCV 2019 · 被引用 189 次
- PanoRecon: Real-Time Panoptic 3D Reconstruction from Monocular VideoDong Wu, Zike Yan, Hongbin ZhaCVPR 2024 · 被引用 8 次
- UnScene3D: Unsupervised 3D Instance Segmentation for Indoor ScenesDávid Rozenberszki, Or Litany, Angela DaiCVPR 2024 · 被引用 25 次
