ROCA: Robust CAD Model Retrieval and Alignment from a Single Image
Can Gümeli, Angela Dai, Matthias Nießner
Abstract
We present ROCA <sup xmlns:mml="http://www.w3.org/1998/Math/MathML" xmlns:xlink="http://www.w3.org/1999/xlink">1</sup> <sup xmlns:mml="http://www.w3.org/1998/Math/MathML" xmlns:xlink="http://www.w3.org/1999/xlink">1</sup> The code is made available at https://github.com/cangurneli/ROCA., a novel end-to-end approach that re-trieves and aligns 3D CAD models from a shape database to a single input image. This enables 3D perception of an ob-served scene from a 2D RGB observation, characterized as a lightweight, compact, clean CAD representation. Core to our approach is our differentiable alignment optimization based on dense 2D-3D object correspondences and Pro-crustes alignment. ROCA can thus provide a robust CAD alignment while simultaneously informing CAD retrieval by leveraging the 2D-3D correspondences to learn geometri-cally similar CAD models. Experiments on challenging, real-world imagery from ScanNet show that ROCA signif-icantly improves on state of the art, from 9.5% to 17.6% in retrieval-aware CAD alignment accuracy.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 83e703d6-c534-427a-a887-7c99c0cd998aCited by top-tier papers31
- PartCrafter: Structured 3D Mesh Generation via Compositional Latent Diffusion TransformersYuchen Lin, Chenguo Lin, Panwang Pan, Honglei Yan et al.NeurIPS 2025 · 89 citations
- CAST: Component-Aligned 3D Scene Reconstruction from an RGB ImageKaixin Yao, Longwen Zhang, Xinhao Yan, Yan Zeng et al.SIGGRAPH 2025 · 30 citations
- DiffCAD: Weakly-Supervised Probabilistic CAD Model Retrieval and Alignment from an RGB ImageDaoyi Gao, Dávid Rozenberszki, Stefan Leutenegger, Angela DaiSIGGRAPH 2024 · 28 citations
- LiteReality: Graphics-Ready 3D Scene Reconstruction from RGB-D ScansZhening Huang, Xiaoyang Wu, Fangcheng Zhong, Hengshuang Zhao et al.NeurIPS 2025 · 26 citations
- WorldGen: From Text to Traversable and Interactive 3D WorldsDilin Wang, Hyunyoung Jung, Tom Monnier, Kihyuk Sohn et al.CVPR 2026 · 24 citations
Builds on10
- CenterNet: Keypoint Triplets for Object DetectionKaiwen Duan, Song Bai, Lingxi Xie, Honggang Qi et al.ICCV 2019 · 3,348 citations
- End-to-End CAD Model Retrieval and 9DoF Alignment in 3D ScansArmen Avetisyan, Angela Dai, Matthias NießnerICCV 2019 · 88 citations
- Neural Non-Rigid TrackingAljaz Bozic, Pablo R. Palafox, Michael Zollhöfer, Angela Dai et al.NeurIPS 2020 · 70 citations
- 3D-RelNet: Joint Object and Relational Network for 3D PredictionNilesh Kulkarni, Ishan Misra, Shubham Tulsiani, Abhinav GuptaICCV 2019 · 48 citations
- Patch2CAD: Patchwise Embedding Learning for In-the-Wild Shape Retrieval from a Single ImageWeicheng Kuo, Anelia Angelova, Tsung-Yi Lin, Angela DaiICCV 2021 · 42 citations
Related papers
- Zero-Shot Inexact CAD Model Alignment from a Single ImagePattaramanee Arsomngern, Sasikarn Khwanmuang, Matthias Nießner, Supasorn SuwajanakornICCV 2025 · 2 citations
- From Points to Multi-Object 3D ReconstructionFrancis Engelmann, Konstantinos Rematas, Bastian Leibe, Vittorio FerrariCVPR 2021
- U-RED: Unsupervised 3D Shape Retrieval and Deformation for Partial Point CloudsYan Di, Chenyangguang Zhang, Ruida Zhang, Fabian Manhardt et al.ICCV 2023 · 15 citations
- Co-op: Correspondence-based Novel Object Pose EstimationSungphill Moon, Hyeontae Son, Dongcheol Hur, Sangwook KimCVPR 2025
- Joint Embedding of 3D Scan and CAD ObjectsManuel Dahnert, Angela Dai, Leonidas J. Guibas, Matthias NießnerICCV 2019 · 36 citations
