VoCAPTER: Voting-based Pose Tracking for Category-level Articulated Object via Inter-frame Priors
Li Zhang, Zean Han, Yan Zhong, Qiaojun Yu, Xingyu Wu, Xue Wang, Rujing Wang
摘要
Articulated objects are common in our daily life. However, current category-level articulation pose works mostly focus on predicting 9D poses on statistical point cloud observations. In this paper, we deal with the problem of category-level online robust 9D pose tracking of articulated objects, where we propose VoCAPTER, a novel 3D Voting-based Category-level Articulated object Pose TrackER. Our VoCAPTER efficiently updates poses between adjacent frames by utilizing partial observations from the current frame and the estimated per-part 9D poses from the previous frame. Specifically, by incorporating prior knowledge of continuous motion relationships between frames, we begin by canonicalizing the input point cloud, casting the pose tracking task as an inter-frame pose increment estimation challenge. Subsequently, to obtain a robust pose-tracking algorithm, our main idea is to leverage SE(3)-invariant features during motion. This is achieved through a voting-based articulation tracking algorithm, which identifies keyframes as reference states for accurate pose updating throughout the entire video sequence. We evaluate the performance of VoCAPTER in the synthetic dataset and real-world scenarios, which demonstrates VoCAPTER's generalization ability to diverse and complicated scenes. Through these experiments, we provide evidence of VoCAPTER's superiority and robustness in multi-frame pose tracking of articulated objects. We believe that this work can facilitate the progress of various fields, including robotics, embodied intelligence, and augmented reality. All the codes will be made publicly available.
问问这篇 Paper
问问你的智能体。
Lune 读过与它相关的顶会 Paper,每个回答都会注明依据哪几篇。
引用它的顶会 Paper8
- UltraHR-100K: Enhancing UHR Image Synthesis with A Large-Scale High-Quality DatasetChen Zhao, En Ci, Yunzhe Xu, Tiehan Fan 等NeurIPS 2025 · 被引用 24 次
- LUVE : Latent-Cascaded Ultra-High-Resolution Video Generation with Dual Frequency ExpertsChen Zhao, Jiawei Chen, Hongyu Li, Zhuoliang Kang 等ICML 2026 · 被引用 16 次
- R^2-Art: Category-Level Articulation Pose Estimation from Single RGB Image via Cascade Render StrategyLi Zhang, Haonan Jiang, Yukang Huo, Yan Zhong 等AAAI 2025 · 被引用 6 次
- MIDGArD: Modular Interpretable Diffusion over Graphs for Articulated DesignsQuentin Leboutet, Nina Wiedemann, Zhipeng Cai, Michael Paulitsch 等NeurIPS 2024 · 被引用 2 次
- Exploring Category-level Articulated Object Pose Tracking on SE(3) ManifoldsXianhui Meng, Yukang Huo, Li Zhang, Liu Liu 等AAAI 2026 · 被引用 1 次
相关 Paper
- KPA-Tracker: Towards Robust and Real-Time Category-Level Articulated Object 6D Pose TrackingLiu Liu, Anran Huang, Qi Wu, Dan Guo 等AAAI 2024 · 被引用 7 次
- CAPTRA: CAtegory-level Pose Tracking for Rigid and Articulated Objects from Point CloudsYijia Weng, He Wang, Qiang Zhou, Yuzhe Qin 等ICCV 2021 · 被引用 119 次
- EfficientCAPER: An End-to-End Framework for Fast and Robust Category-Level Articulated Object Pose EstimationXinyi Yu, Haonan Jiang, Li Zhang, Lin Yuanbo Wu 等NeurIPS 2024
- Category-Level Articulated Object 9D Pose Estimation via Reinforcement LearningLiu Liu, Jianming Du, Hao Wu, Xun Yang 等ACM MM 2023 · 被引用 14 次
- Category-Level Articulated Object Pose EstimationXiaolong Li, He Wang, Li Yi, Leonidas J. Guibas 等CVPR 2020
