DVMNet: Computing Relative Pose for Unseen Objects Beyond Hypotheses
Chen Zhao, Tong Zhang, Zheng Dang, Mathieu Salzmann
摘要
Determining the relative pose of an object between two images is pivotal to the success of generalizable object pose estimation. Existing approaches typically approximate the continuous pose representation with a large number of discrete pose hypotheses, which incurs a computationally expensive process of scoring each hypothesis at test time. By contrast, we present a Deep Voxel Matching Network (DVMNet) that eliminates the need for pose hypotheses and computes the relative object pose in a single pass. To this end, we map the two input RGB images, reference and query, to their respective voxelized 3D representations. We then pass the resulting voxels through a pose estimation module, where the voxels are aligned and the pose is computed in an end-to-end fashion by solving a least-squares problem. To enhance robustness, we introduce a weighted closest voxel algorithm capable of mitigating the impact of noisy voxels. We conduct extensive experiments on the CO3D, LINEMOD, and Objaverse datasets, demonstrating that our method delivers more accurate relative pose estimates for novel objects at a lower computational cost compared to state-of-the-art methods. Our code is released at: https://github.com/sailor-z/DVMNet/ .
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了最后一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper4
- BoxDreamer: Dreaming Box Corners for Generalizable Object Pose EstimationYuanhong Yu, Xingyi He, Chen Zhao, Junhao Yu 等ICCV 2025 · 被引用 3 次
- SingRef6D: Monocular Novel Object Pose Estimation with a Single RGB ReferenceJiahui Wang, Haiyue Zhu, Haoren Guo, Abdullah Al Mamun 等NeurIPS 2025 · 被引用 2 次
- UNOPose: Unseen Object Pose Estimation with an Unposed RGB-D Reference ImageXingyu Liu, Gu Wang, Ruida Zhang, Chenyangguang Zhang 等CVPR 2025
- Structure-Aware Correspondence Learning for Relative Pose EstimationYihan Chen, Wenfei Yang, Huan Ren, Shifeng Zhang 等CVPR 2025
它引用的顶会 Paper18
- An Image is Worth 16x16 Words: Transformers for Image Recognition at ScaleAlexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn 等ICLR 2021 · 被引用 21,477 次
- Zero-1-to-3: Zero-shot One Image to 3D ObjectRuoshi Liu, Rundi Wu, Basile Van Hoorick, Pavel Tokmakov 等ICCV 2023 · 被引用 1,662 次
- Common Objects in 3D: Large-Scale Learning and Evaluation of Real-life 3D Category ReconstructionJeremy Reizenstein, Roman Shapovalov, Philipp Henzler, Luca Sbordone 等ICCV 2021 · 被引用 686 次
- Learning Two-View Correspondences and Geometry Using Order-Aware NetworkJiahui Zhang, Dawei Sun, Zixin Luo, Anbang Yao 等ICCV 2019 · 被引用 362 次
- OnePose++: Keypoint-Free One-Shot Object Pose Estimation without CAD ModelsXingyi He, Jiaming Sun, Yuang Wang, Di Huang 等NeurIPS 2022 · 被引用 190 次
相关 Paper
- 3D-Aware Hypothesis & Verification for Generalizable Relative Object Pose EstimationChen Zhao, Tong Zhang, Mathieu SalzmannICLR 2024 · 被引用 13 次
- DCNet: Dense Correspondence Neural Network for 6DoF Object Pose Estimation in Occluded ScenesZhi Chen, Wei Yang, Zhenbo Xu, Xike Xie 等ACM MM 2020 · 被引用 3 次
- GDR-Net: Geometry-Guided Direct Regression Network for Monocular 6D Object Pose EstimationGu Wang, Fabian Manhardt, Federico Tombari, Xiangyang JiCVPR 2021
- PVN3D: A Deep Point-Wise 3D Keypoints Voting Network for 6DoF Pose EstimationYisheng He, Wei Sun, Haibin Huang, Jianran Liu 等CVPR 2020
- CDPN: Coordinates-Based Disentangled Pose Network for Real-Time RGB-Based 6-DoF Object Pose EstimationZhigang Li, Gu Wang, Xiangyang JiICCV 2019 · 被引用 482 次
