Learning Local RGB-to-CAD Correspondences for Object Pose Estimation
Georgios Georgakis, Srikrishna Karanam, Ziyan Wu, Jana Kosecka
Abstract
We consider the problem of 3D object pose estimation. While much recent work has focused on the RGB domain, the reliance on accurately annotated images limits generalizability and scalability. On the other hand, the easily available object CAD models are rich sources of data, providing a large number of synthetically rendered images. In this paper, we solve this key problem of existing methods requiring expensive 3D pose annotations by proposing a new method that matches RGB images to CAD models for object pose estimation. Our key innovations compared to existing work include removing the need for either real-world textures for CAD models or explicit 3D pose annotations for RGB images. We achieve this through a series of objectives that learn how to select keypoints and enforce viewpoint and modality invariance across RGB images and CAD model renderings. Our experiments demonstrate that the proposed method can reliably estimate object pose in RGB images and generalize to object instances not seen during training.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 113a950b-63d7-4215-99b8-6f7add8babe1Cited by top-tier papers5
- Patch2CAD: Patchwise Embedding Learning for In-the-Wild Shape Retrieval from a Single ImageWeicheng Kuo, Anelia Angelova, Tsung-Yi Lin, Angela DaiICCV 2021 · 42 citations
- SD-Pose: Semantic Decomposition for Cross-Domain 6D Object Pose EstimationZhigang Li, Yinlin Hu, Mathieu Salzmann, Xiangyang JiAAAI 2021 · 16 citations
- Part-Based Models Improve Adversarial RobustnessChawin Sitawarin, Kornrapat Pongmala, Yizheng Chen, Nicholas Carlini et al.ICLR 2023 · 2 citations
- Visual Localization using Imperfect 3D Models from the InternetVojtech Panek, Zuzana Kukelova, Torsten SattlerCVPR 2023
- Learning Canonical Shape Space for Category-Level 6D Object Pose and Size EstimationDengsheng Chen, Jun Li, Zheng Wang, Kai XuCVPR 2020
Related papers
- Learning Deep Network for Detecting 3D Object Keypoints and 6D PosesWanqing Zhao, Shaobo Zhang, Ziyu Guan, Wei Zhao et al.CVPR 2020
- ONDA-Pose: Occlusion-Aware Neural Domain Adaptation for Self-Supervised 6D Object Pose EstimationTao Tan, Qiulei DongCVPR 2025
- Templates for 3D Object Pose Estimation Revisited: Generalization to New Objects and Robustness to OcclusionsVan Nguyen Nguyen, Yinlin Hu, Yang Xiao, Mathieu Salzmann et al.CVPR 2022 · 84 citations
- DiffCAD: Weakly-Supervised Probabilistic CAD Model Retrieval and Alignment from an RGB ImageDaoyi Gao, Dávid Rozenberszki, Stefan Leutenegger, Angela DaiSIGGRAPH 2024 · 28 citations
- Reconstruct Locally, Localize Globally: A Model Free Method for Object Pose EstimationMing Cai, Ian ReidCVPR 2020
