ELLIPSDF: Joint Object Pose and Shape Optimization with a Bi-level Ellipsoid and Signed Distance Function Description
Mo Shan, Qiaojun Feng, You-Yi Jau, Nikolay Atanasov
Abstract
Autonomous systems need to understand the semantics and geometry of their surroundings in order to comprehend and safely execute object-level task specifications. This paper proposes an expressive yet compact model for joint object pose and shape optimization, and an associated optimization algorithm to infer an object-level map from multi-view RGB-D camera observations. The model is expressive because it captures the identities, positions, orientations, and shapes of objects in the environment. It is compact because it relies on a low-dimensional latent representation of implicit object shape, allowing onboard storage of large multi-category object maps. Different from other works that rely on a single object representation format, our approach has a bi-level object model that captures both the coarse level scale as well as the fine level shape details. Our approach is evaluated on the large-scale real-world ScanNet dataset and compared against state-of-the-art methods.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext ceea729a-6b68-4121-98fa-9d707f8bc37aCited by top-tier papers2
- Living Scenes: Multi-object Relocalization and Reconstruction in Changing 3D EnvironmentsLiyuan Zhu, Shengyu Huang, Konrad Schindler, Iro ArmeniCVPR 2024 · 10 citations
- O^2-Recon: Completing 3D Reconstruction of Occluded Objects in the Scene with a Pre-trained 2D Diffusion ModelYubin Hu, Sheng Ye, Wang Zhao, Matthieu Lin et al.AAAI 2024 · 7 citations
Builds on4
- End-to-End CAD Model Retrieval and 9DoF Alignment in 3D ScansArmen Avetisyan, Angela Dai, Matthias NießnerICCV 2019 · 88 citations
- Learning to Group: A Bottom-Up Framework for 3D Part Discovery in Unseen CategoriesTiange Luo, Kaichun Mo, Zhiao Huang, Jiarui Xu et al.ICLR 2020 · 44 citations
- DualSDF: Semantic Shape Manipulation Using a Two-Level RepresentationZekun Hao, Hadar Averbuch-Elor, Noah Snavely, Serge J. BelongieCVPR 2020
- FroDO: From Detections to 3D ObjectsMartin Rünz, Kejie Li, Meng Tang, Lingni Ma et al.CVPR 2020
Related papers
- Learning 3D Scene Priors with 2D SupervisionYinyu Nie, Angela Dai, Xiaoguang Han, Matthias NießnerCVPR 2023
- Rooms from Motion: Un-posed Indoor 3D Object Detection as Localization and MappingJustin Lazarow, Kai Kang, Afshin DehghanNeurIPS 2025 · 2 citations
- Holistic 3D Scene Understanding From a Single Image With Implicit RepresentationCheng Zhang, Zhaopeng Cui, Yinda Zhang, Bing Zeng et al.CVPR 2021
- ODAM: Object Detection, Association, and Mapping using Posed RGB VideoKejie Li, Daniel DeTone, Steven Chen, Minh Vo et al.ICCV 2021 · 31 citations
- Patch2CAD: Patchwise Embedding Learning for In-the-Wild Shape Retrieval from a Single ImageWeicheng Kuo, Anelia Angelova, Tsung-Yi Lin, Angela DaiICCV 2021 · 42 citations
