DSC-PoseNet: Learning 6DoF Object Pose Estimation via Dual-Scale Consistency
Zongxin Yang, Xin Yu, Yi Yang
摘要
Compared to 2D object bounding-box labeling, it is very difficult for humans to annotate 3D object poses, especially when depth images of scenes are unavailable. This paper investigates whether we can estimate the object poses effectively when only RGB images and 2D object annotations are given. To this end, we present a two-step pose estimation framework to attain 6DoF object poses from 2D object bounding-boxes. In the first step, the framework learns to segment objects from real and synthetic data in a weaklysupervised fashion, and the segmentation masks will act as a prior for pose estimation. In the second step, we design a dual-scale pose estimation network, namely DSC-PoseNet, to predict object poses by employing a differential renderer. To be specific, our DSC-PoseNet firstly predicts object poses in the original image scale by comparing the segmentation masks and the rendered visible object masks. Then, we resize object regions to a fixed scale to estimate poses once again. In this fashion, we eliminate large scale variations and focus on rotation estimation, thus facilitating pose estimation. Moreover, we exploit the initial pose estimation to generate pseudo ground-truth to train our DSC-PoseNet in a self-supervised manner. The estimation results in these two scales are ensembled as our final pose estimation. Extensive experiments on widely-used benchmarks demonstrate that our method outperforms state-of-the-art models trained on synthetic data by a large margin and even is on par with several fully-supervised methods.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper7
- Self-Supervised Category-Level 6D Object Pose Estimation with Deep Implicit Shape RepresentationWanli Peng, Jianhang Yan, Hongtao Wen, Yi SunAAAI 2022 · 被引用 46 次
- JOTR: 3D Joint Contrastive Learning with Transformers for Occluded Human Mesh RecoveryJiahao Li, Zongxin Yang, Xiaohan Wang, Jianxin Ma 等ICCV 2023 · 被引用 22 次
- Pseudo Flow Consistency for Self-Supervised 6D Object Pose EstimationYang Hai, Rui Song, Jiaojiao Li, David Ferstl 等ICCV 2023 · 被引用 13 次
- Environment-Agnostic Pose: Generating Environment-Independent Object Representations for 6D Pose EstimationShaobo Zhang, Yuhang Huang, Wanqing Zhao, Wei Zhao 等ICCV 2025 · 被引用 3 次
- ONDA-Pose: Occlusion-Aware Neural Domain Adaptation for Self-Supervised 6D Object Pose EstimationTao Tan, Qiulei DongCVPR 2025
它引用的顶会 Paper9
- Pix2Pose: Pixel-Wise Coordinate Regression of Objects for 6D Pose EstimationKiru Park, Timothy Patten, Markus VinczeICCV 2019 · 被引用 527 次
- CDPN: Coordinates-Based Disentangled Pose Network for Real-Time RGB-Based 6-DoF Object Pose EstimationZhigang Li, Gu Wang, Xiangyang JiICCV 2019 · 被引用 482 次
- Explaining the Ambiguity of Object Detection and 6D Pose From Visual DataFabian Manhardt, Diego Martín Arroyo, Christian Rupprecht, Benjamin Busam 等ICCV 2019 · 被引用 139 次
- Weakly-Supervised Salient Object Detection via Scribble AnnotationsJing Zhang, Xin Yu, Aixuan Li, Peipei Song 等CVPR 2020
- Single-Stage 6D Object Pose EstimationYinlin Hu, Pascal Fua, Wei Wang, Mathieu SalzmannCVPR 2020
相关 Paper
- SMOC-Net: Leveraging Camera Pose for Self-Supervised Monocular Object Pose EstimationTao Tan, Qiulei DongCVPR 2023
- PFRL: Pose-Free Reinforcement Learning for 6D Pose EstimationJianzhun Shao, Yuhang Jiang, Gu Wang, Zhigang Li 等CVPR 2020
- Learning Deep Network for Detecting 3D Object Keypoints and 6D PosesWanqing Zhao, Shaobo Zhang, Ziyu Guan, Wei Zhao 等CVPR 2020
- Geometry-Driven Self-Supervised Method for 3D Human Pose EstimationYang Li, Kan Li, Shuai Jiang, Ziyue Zhang 等AAAI 2020 · 被引用 40 次
- G2L-Net: Global to Local Network for Real-Time 6D Pose Estimation With Embedding Vector FeaturesWei Chen, Xi Jia, Hyung Jin Chang, Jinming Duan 等CVPR 2020
