3D-Aware Hypothesis & Verification for Generalizable Relative Object Pose Estimation
Chen Zhao, Tong Zhang, Mathieu Salzmann
摘要
Prior methods that tackle the problem of generalizable object pose estimation highly rely on having dense views of the unseen object. By contrast, we address the scenario where only a single reference view of the object is available. Our goal then is to estimate the relative object pose between this reference view and a query image that depicts the object in a different pose. In this scenario, robust generalization is imperative due to the presence of unseen objects during testing and the large-scale object pose variation between the reference and the query. To this end, we present a new hypothesis-and-verification framework, in which we generate and evaluate multiple pose hypotheses, ultimately selecting the most reliable one as the relative object pose. To measure reliability, we introduce a 3D-aware verification that explicitly applies 3D transformations to the 3D object representations learned from the two input images. Our comprehensive experiments on the Objaverse, LINEMOD, and CO3D datasets evidence the superior accuracy of our approach in relative pose estimation and its robustness in large-scale pose variations, when dealing with unseen objects.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper8
- Sparse-view Pose Estimation and Reconstruction via Analysis by Generative SynthesisQitao Zhao, Shubham TulsianiNeurIPS 2024 · 被引用 10 次
- DVMNet: Computing Relative Pose for Unseen Objects Beyond HypothesesChen Zhao, Tong Zhang, Zheng Dang, Mathieu SalzmannCVPR 2024 · 被引用 5 次
- Vision Foundation Model Enables Generalizable Object Pose EstimationKai Chen, Yiyao Ma, Xingyu Lin, Stephen James 等NeurIPS 2024 · 被引用 5 次
- BoxDreamer: Dreaming Box Corners for Generalizable Object Pose EstimationYuanhong Yu, Xingyi He, Chen Zhao, Junhao Yu 等ICCV 2025 · 被引用 3 次
- SingRef6D: Monocular Novel Object Pose Estimation with a Single RGB ReferenceJiahui Wang, Haiyue Zhu, Haoren Guo, Abdullah Al Mamun 等NeurIPS 2025 · 被引用 2 次
它引用的顶会 Paper20
- A Simple Framework for Contrastive Learning of Visual RepresentationsTing Chen, Simon Kornblith, Mohammad Norouzi, Geoffrey E. HintonICML 2020 · 被引用 24,064 次
- An Image is Worth 16x16 Words: Transformers for Image Recognition at ScaleAlexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn 等ICLR 2021 · 被引用 21,477 次
- SimMIM: a Simple Framework for Masked Image ModelingZhenda Xie, Zheng Zhang, Yue Cao, Yutong Lin 等CVPR 2022 · 被引用 1,129 次
- Common Objects in 3D: Large-Scale Learning and Evaluation of Real-life 3D Category ReconstructionJeremy Reizenstein, Roman Shapovalov, Philipp Henzler, Luca Sbordone 等ICCV 2021 · 被引用 686 次
- OnePose++: Keypoint-Free One-Shot Object Pose Estimation without CAD ModelsXingyi He, Jiaming Sun, Yuang Wang, Di Huang 等NeurIPS 2022 · 被引用 190 次
相关 Paper
- Structure-Aware Correspondence Learning for Relative Pose EstimationYihan Chen, Wenfei Yang, Huan Ren, Shifeng Zhang 等CVPR 2025
- UNOPose: Unseen Object Pose Estimation with an Unposed RGB-D Reference ImageXingyu Liu, Gu Wang, Ruida Zhang, Chenyangguang Zhang 等CVPR 2025
- One2Any: One-Reference 6D Pose Estimation for Any ObjectMengya Liu, Siyuan Li, Ajad Chhatkuli, Prune Truong 等CVPR 2025
- PoseGAM: Robust Unseen Object Pose Estimation via Geometry-Aware Multi-View ReasoningJianqi Chen, Biao Zhang, Xiangjun Tang, Peter WonkaCVPR 2026
- CoordAR: One-Reference 6D Pose Estimation of Novel Objects via Autoregressive Coordinate Map GenerationDexin Zuo, Ang Li, Wei Wang, Wenxian Yu 等AAAI 2026
