Pseudo Flow Consistency for Self-Supervised 6D Object Pose Estimation
Yang Hai, Rui Song, Jiaojiao Li, David Ferstl, Yinlin Hu
Abstract
Most self-supervised 6D object pose estimation methods can only work with additional depth information or rely on the accurate annotation of 2D segmentation masks, limiting their application range. In this paper, we propose a 6D object pose estimation method that can be trained with pure RGB images without any auxiliary information. We first obtain a rough pose initialization from networks trained on synthetic images rendered from the target’s 3D mesh. Then, we introduce a refinement strategy leveraging the geometry constraint in synthetic-to-real image pairs from multiple different views. We formulate this geometry constraint as pixel-level flow consistency between the training images with dynamically generated pseudo labels. We evaluate our method on three challenging datasets and demonstrate that it outperforms state-of-the-art self-supervised methods significantly, with neither 2D annotations nor additional depth images.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 51ecfbe8-b7e9-4b71-b85b-73dd7df85674Cited by top-tier papers3
- SCFlow2: Plug-and-Play Object Pose Refiner with Shape-Constraint Scene FlowQingyuan Wang, Rui Song, Jiaojiao Li, Kerui Cheng et al.CVPR 2025
- Hierarchical Flow Diffusion for Efficient Frame InterpolationYang Hai, Guo Wang, Tan Su, Wenjie Jiang et al.CVPR 2025
- Ultra-Precision 6DoF Pose Estimation Using 2-D Interpolated Discrete Fourier TransformGuowei Shi, Zian Mao, Peisen HuangICCV 2025
Builds on27
- FixMatch: Simplifying Semi-Supervised Learning with Consistency and ConfidenceKihyuk Sohn, David Berthelot, Nicholas Carlini, Zizhao Zhang et al.NeurIPS 2020 · 5,129 citations
- FlexMatch: Boosting Semi-Supervised Learning with Curriculum Pseudo LabelingBowen Zhang, Yidong Wang, Wenxin Hou, Hao Wu et al.NeurIPS 2021 · 1,389 citations
- Soft Rasterizer: A Differentiable Renderer for Image-Based 3D ReasoningShichen Liu, Weikai Chen, Tianye Li, Hao LiICCV 2019 · 789 citations
- End-to-End Semi-Supervised Object Detection with Soft TeacherMengde Xu, Zheng Zhang, Han Hu, Jianfeng Wang et al.ICCV 2021 · 622 citations
- Pix2Pose: Pixel-Wise Coordinate Regression of Objects for 6D Pose EstimationKiru Park, Timothy Patten, Markus VinczeICCV 2019 · 527 citations
Related papers
- DSC-PoseNet: Learning 6DoF Object Pose Estimation via Dual-Scale ConsistencyZongxin Yang, Xin Yu, Yi YangCVPR 2021
- PFRL: Pose-Free Reinforcement Learning for 6D Pose EstimationJianzhun Shao, Yuhang Jiang, Gu Wang, Zhigang Li et al.CVPR 2020
- SMOC-Net: Leveraging Camera Pose for Self-Supervised Monocular Object Pose EstimationTao Tan, Qiulei DongCVPR 2023
- Learning Deep Network for Detecting 3D Object Keypoints and 6D PosesWanqing Zhao, Shaobo Zhang, Ziyu Guan, Wei Zhao et al.CVPR 2020
- Self-Supervised Geometric Correspondence for Category-Level 6D Object Pose Estimation in the WildKaifeng Zhang, Yang Fu, Shubhankar Borse, Hong Cai et al.ICLR 2023 · 8 citations
