R^2-Art: Category-Level Articulation Pose Estimation from Single RGB Image via Cascade Render Strategy
Li Zhang, Haonan Jiang, Yukang Huo, Yan Zhong, Jianan Wang, Xue Wang, Rujing Wang, Liu Liu
Abstract
Human life is filled with articulated objects. Previous works for estimating the pose of category-level articulated objects rely on costly 3D point clouds or RGB-D images. In this paper, our goal is to estimate category-level articulation poses from a single RGB image, where we propose R 2 -Art, a novel category-level Articulation pose estimation framework from a single RGB image and a cascade Render strategy. Given an RGB image as input, R 2 -Art estimates per-part 6D pose for the articulation. Specifically, we design parallel regression branches tailored to generate camera-to-root translation and rotation. Using the predicted joint states, we perform PC prior transformation and deformation with a joint-centric modeling approach. For further refinement, a cascade render strategy is proposed for projecting the 3D deformed prior onto the 2D mask. Extensive experiments are provided to validate our R 2 -Art on various datasets ranging from synthetic datasets to real-world scenarios, demonstrating the superior performance and robustness of the R 2 -Art. We believe that this work has the potential to be applied in many fields including robotics, embodied intelligence, and augmented reality.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 3b407393-0807-47d2-a619-71df2b76ed4aCited by top-tier papers4
- LUVE : Latent-Cascaded Ultra-High-Resolution Video Generation with Dual Frequency ExpertsChen Zhao, Jiawei Chen, Hongyu Li, Zhuoliang Kang et al.ICML 2026 · 16 citations
- DICArt: Advancing Category-level Articulated Object Pose Estimation in Discrete State-SpacesLi Zhang, Mingyu Mei, Ailing Wang, Xianhui Meng et al.CVPR 2026 · 2 citations
- DFGAP: Towards Depth-Free Cross-Category GAParts Perception via Uncertainty-Quantified ModelingXueyu Yuan, Jiarui Zhang, Jiangqi Song, Liu Liu et al.ACM MM 2025
- Generalizable and Actionable Parts Pose Estimation with Symmetry Annotation-Free Learning Strategywenxiao chen, Xueyu Yuan, Liu Liu, Di Wu et al.ICML 2026
Builds on10
- Camera Distance-Aware Top-Down Approach for 3D Multi-Person Pose Estimation From a Single RGB ImageGyeongsik Moon, Ju Yong Chang, Kyoung Mu LeeICCV 2019 · 368 citations
- Where2Act: From Pixels to Actions for Articulated 3D ObjectsKaichun Mo, Leonidas J. Guibas, Mustafa Mukadam, Abhinav Gupta et al.ICCV 2021 · 240 citations
- GPV-Pose: Category-level Object Pose Estimation via Geometry-guided Point-wise VotingYan Di, Ruida Zhang, Zhiqiang Lou, Fabian Manhardt et al.CVPR 2022 · 141 citations
- CAPTRA: CAtegory-level Pose Tracking for Rigid and Articulated Objects from Point CloudsYijia Weng, He Wang, Qiang Zhou, Yuzhe Qin et al.ICCV 2021 · 119 citations
- End-to-End CAD Model Retrieval and 9DoF Alignment in 3D ScansArmen Avetisyan, Angela Dai, Matthias NießnerICCV 2019 · 88 citations
Related papers
- KPA-Tracker: Towards Robust and Real-Time Category-Level Articulated Object 6D Pose TrackingLiu Liu, Anran Huang, Qi Wu, Dan Guo et al.AAAI 2024 · 7 citations
- Category-Level Articulated Object 9D Pose Estimation via Reinforcement LearningLiu Liu, Jianming Du, Hao Wu, Xun Yang et al.ACM MM 2023 · 14 citations
- EfficientCAPER: An End-to-End Framework for Fast and Robust Category-Level Articulated Object Pose EstimationXinyi Yu, Haonan Jiang, Li Zhang, Lin Yuanbo Wu et al.NeurIPS 2024
- CAP-Net: A Unified Network for 6D Pose and Size Estimation of Categorical Articulated Parts from a Single RGB-D ImageJingshun Huang, Haitao Lin, Tianyu Wang, Yanwei Fu et al.CVPR 2025
- Category-Level Articulated Object Pose EstimationXiaolong Li, He Wang, Li Yi, Leonidas J. Guibas et al.CVPR 2020
