Single-View Robot Pose and Joint Angle Estimation via Render & Compare
Yann Labbé, Justin Carpentier, Mathieu Aubry, Josef Sivic
Abstract
We introduce RoboPose, a method to estimate the joint angles and the 6D camera-to-robot pose of a known articulated robot from a single RGB image. This is an important problem to grant mobile and itinerant autonomous systems the ability to interact with other robots using only visual information in non-instrumented environments, especially in the context of collaborative robotics. It is also challenging because robots have many degrees of freedom and an infinite space of possible configurations that often result in self-occlusions and depth ambiguities when imaged by a single camera. The contributions of this work are three-fold. First, we introduce a new render & compare approach for estimating the 6D pose and joint angles of an articulated robot that can be trained from synthetic data, generalizes to new unseen robot configurations at test time, and can be applied to a variety of robots. Second, we experimentally demonstrate the importance of the robot parametrization for the iterative pose updates and design a parametrization strategy that is independent of the robot structure. Finally, we show experimental results on existing benchmark datasets for four different robots and demonstrate that our method significantly outperforms the state of the art. Code and pre-trained models are available on the project webpage [1].
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext dbfb2ee7-84f0-4c84-8184-f90146f2e57fCited by top-tier papers10
- Pooling Revisited: Your Receptive Field is SuboptimalDong-Hwan Jang, Sanghyeok Chu, Joonhyuk Kim, Bohyung HanCVPR 2022 · 13 citations
- Self-Supervised Category-Level Articulated Object Pose Estimation with Part-Level SE(3) EquivarianceXueyi Liu, Ji Zhang, Ruizhen Hu, Haibin Huang et al.ICLR 2023 · 3 citations
- Know Thyself: Transferable Visual Control Policies Through Robot-AwarenessEdward S. Hu, Kun Huang, Oleh Rybkin, Dinesh JayaramanICLR 2022 · 2 citations
- EgoRoC: Towards Egocentric Robotic Control via Task-Agnostic Visual AlignmentWei Feng, Chi Zhang, Nan Li, Qian Zhang et al.CVPR 2026
- Category-Level Articulated Object Pose EstimationXiaolong Li, He Wang, Li Yi, Leonidas J. Guibas et al.CVPR 2020
Builds on5
- Pix2Pose: Pixel-Wise Coordinate Regression of Objects for 6D Pose EstimationKiru Park, Timothy Patten, Markus VinczeICCV 2019 · 527 citations
- Disentangling Monocular 3D Object DetectionAndrea Simonelli, Samuel Rota Bulò, Lorenzo Porzi, Manuel Lopez-Antequera et al.ICCV 2019 · 504 citations
- DPOD: 6D Pose Object Detector and RefinerSergey Zakharov, Ivan Shugurov, Slobodan IlicICCV 2019 · 486 citations
- Category-Level Articulated Object Pose EstimationXiaolong Li, He Wang, Li Yi, Leonidas J. Guibas et al.CVPR 2020
- HybridPose: 6D Object Pose Estimation Under Hybrid RepresentationsChen Song, Jiaru Song, Qixing HuangCVPR 2020
Related papers
- Robot Structure Prior Guided Temporal Attention for Camera-to-Robot Pose Estimation from Image SequenceYang Tian, Jiyao Zhang, Zekai Yin, Hao DongCVPR 2023
- RoboPEPP: Vision-Based Robot Pose and Joint Angle Estimation through Embedding Predictive Pre-TrainingRaktim Gautam Goswami, Prashanth Krishnamurthy, Yann LeCun, Farshad KhorramiCVPR 2025
- Any6D: Model-free 6D Pose Estimation of Novel ObjectsTaeyeop Lee, Bowen Wen, Minjun Kang, Gyuree Kang et al.CVPR 2025
- CDPN: Coordinates-Based Disentangled Pose Network for Real-Time RGB-Based 6-DoF Object Pose EstimationZhigang Li, Gu Wang, Xiangyang JiICCV 2019 · 482 citations
- AlignPose: Generalizable 6D Pose Estimation via Multi-view Feature-metric AlignmentAnna Sárová Mikestíková, Médéric Fourmy, Martin Cífka, Josef Sivic et al.CVPR 2026
