Self-Supervised Geometric Correspondence for Category-Level 6D Object Pose Estimation in the Wild
Kaifeng Zhang, Yang Fu, Shubhankar Borse, Hong Cai, Fatih Porikli, Xiaolong Wang
摘要
While 6D object pose estimation has wide applications across computer vision and robotics, it remains far from being solved due to the lack of annotations. The problem becomes even more challenging when moving to category-level 6D pose, which requires generalization to unseen instances. Current approaches are restricted by leveraging annotations from simulation or collected from humans. In this paper, we overcome this barrier by introducing a self-supervised learning approach trained directly on large-scale real-world object videos for category-level 6D pose estimation in the wild. Our framework reconstructs the canonical 3D shape of an object category and learns dense correspondences between input images and the canonical shape via surface embedding. For training, we propose novel geometrical cycle-consistency losses which construct cycles across 2D-3D spaces, across different instances and different time steps. The learned correspondence can be applied for 6D pose estimation and other downstream tasks such as keypoint transfer. Surprisingly, our method, without any human annotations or simulators, can achieve on-par or even better performance than previous supervised or semisupervised methods on in-the-wild images. Code and videos are available at https://kywind.github.io/self-pose .
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了最后一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper11
- IST-Net: Prior-free Category-level Pose Estimation with Implicit Space TransformationJianhui Liu, Yukang Chen, Xiaoqing Ye, Xiaojuan QiICCV 2023 · 被引用 64 次
- A General Protocol to Probe Large Vision Models for 3D Physical UnderstandingGuanqi Zhan, Chuanxia Zheng, Weidi Xie, Andrew ZissermanNeurIPS 2024 · 被引用 37 次
- RGBD Objects in the Wild: Scaling Real-World 3D Object Learning from RGB-D VideosHongchi Xia, Yang Fu, Sifei Liu, Xiaolong WangCVPR 2024 · 被引用 14 次
- Source-Free and Image-Only Unsupervised Domain Adaptation for Category Level Object Pose EstimationPrakhar Kaushik, Aayush Mishra, Adam Kortylewski, Alan L. YuilleICLR 2024 · 被引用 10 次
- 3D Neural Embedding Likelihood: Probabilistic Inverse Graphics for Robust 6D Pose EstimationGuangyao Zhou, Nishad Gothoskar, Lirui Wang, Joshua B. Tenenbaum 等ICCV 2023 · 被引用 5 次
它引用的顶会 Paper22
- Emerging Properties in Self-Supervised Vision TransformersMathilde Caron, Hugo Touvron, Ishan Misra, Hervé Jégou 等ICCV 2021 · 被引用 8,921 次
- Soft Rasterizer: A Differentiable Renderer for Image-Based 3D ReasoningShichen Liu, Weikai Chen, Tianye Li, Hao LiICCV 2019 · 被引用 789 次
- Space-Time Correspondence as a Contrastive Random WalkAllan Jabri, Andrew Owens, Alexei A. EfrosNeurIPS 2020 · 被引用 356 次
- DualPoseNet: Category-level 6D Object Pose and Size Estimation Using Dual Pose Network with Refined Learning of Pose ConsistencyJiehong Lin, Zewei Wei, Zhihao Li, Songcen Xu 等ICCV 2021 · 被引用 169 次
- GPV-Pose: Category-level Object Pose Estimation via Geometry-guided Point-wise VotingYan Di, Ruida Zhang, Zhiqiang Lou, Fabian Manhardt 等CVPR 2022 · 被引用 141 次
相关 Paper
- Category-Level 6D Object Pose Estimation in the Wild: A Semi-Supervised Learning Approach and A New DatasetYanjie Ze, Xiaolong WangNeurIPS 2022 · 被引用 104 次
- Unsupervised Learning of Category-Level 3D Pose from Object-Centric VideosLeonhard Sommer, Artur Jesslen, Eddy Ilg, Adam KortylewskiCVPR 2024
- Self-Supervised Category-Level 6D Object Pose Estimation with Deep Implicit Shape RepresentationWanli Peng, Jianhang Yan, Hongtao Wen, Yi SunAAAI 2022 · 被引用 46 次
- Leveraging SE(3) Equivariance for Self-supervised Category-Level Object Pose Estimation from Point CloudsXiaolong Li, Yijia Weng, Li Yi, Leonidas J. Guibas 等NeurIPS 2021 · 被引用 61 次
- Instance-Adaptive and Geometric-Aware Keypoint Learning for Category-Level 6D Object Pose EstimationXiao Lin, Wenfei Yang, Yuan Gao, Tianzhu ZhangCVPR 2024
