Key-Grid: Unsupervised 3D Keypoints Detection using Grid Heatmap Features
Chengkai Hou, Zhengrong Xue, Bingyang Zhou, Jinghan Ke, Lin Shao, Huazhe Xu
Abstract
Detecting 3D keypoints with semantic consistency is widely used in many scenarios such as pose estimation, shape registration and robotics. Currently, most unsupervised 3D keypoint detection methods focus on the rigid-body objects. However, when faced with deformable objects, the keypoints they identify do not preserve semantic consistency well. In this paper, we introduce an innovative unsupervised keypoint detector Key-Grid for both the rigid-body and deformable objects, which is an autoencoder framework. The encoder predicts keypoints and the decoder utilizes the generated keypoints to reconstruct the objects. Unlike previous work, we leverage the identified keypoint in formation to form a 3D grid feature heatmap called grid heatmap, which is used in the decoder section. Grid heatmap is a novel concept that represents the latent variables for grid points sampled uniformly in the 3D cubic space, where these variables are the shortest distance between the grid points and the skeleton connected by keypoint pairs. Meanwhile, we incorporate the information from each layer of the encoder into the decoder section. We conduct an extensive evaluation of Key-Grid on a list of benchmark datasets. Key-Grid achieves the state-of-the-art performance on the semantic consistency and position accuracy of keypoints. Moreover, we demonstrate the robustness of Key-Grid to noise and downsampling. In addition, we achieve SE-(3) invariance of keypoints though generalizing Key-Grid to a SE(3)-invariant backbone.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 97cb8599-86c8-4d7f-a760-0daa1e7371b0Cited by top-tier papers2
- Topology-aware Feature Propagation for Unsupervised Non-rigid Point Cloud CorrespondenceHaozhe Chen, Rui Li, Zhengbao Wang, Xinhao Zhu et al.CVPR 2026
- Geometry-aware RL for Manipulation of Varying Shapes and Deformable ObjectsTai Hoang, Huy Le, Philipp Becker, Ngo Anh Vien et al.ICLR 2025
Builds on14
- A-NeRF: Articulated Neural Radiance Fields for Learning Human Shape, Appearance, and PoseShih-Yang Su, Frank Yu, Michael Zollhöfer, Helge RhodinNeurIPS 2021 · 316 citations
- USIP: Unsupervised Stable Interest Point Detection From 3D Point CloudsJiaxin Li, Gim Hee LeeICCV 2019 · 206 citations
- LAKe-Net: Topology-Aware Point Cloud Completion by Localizing Aligned KeypointsJunshu Tang, Zhijun Gong, Ran Yi, Yuan Xie et al.CVPR 2022 · 74 citations
- Canonical Capsules: Self-Supervised Capsules in Canonical PoseWeiwei Sun, Andrea Tagliasacchi, Boyang Deng, Sara Sabour et al.NeurIPS 2021 · 44 citations
- UKPGAN: A General Self-Supervised Keypoint DetectorYang You, Wenhai Liu, Yanjie Ze, Yong-Lu Li et al.CVPR 2022 · 30 citations
Related papers
- Skeleton Merger: An Unsupervised Aligned Keypoint DetectorRuoxi Shi, Zhengrong Xue, Yang You, Cewu LuCVPR 2021
- Semi-supervised Keypoint LocalizationOlga Moskvyak, Frédéric Maire, Feras Dayoub, Mahsa BaktashmotlaghICLR 2021 · 17 citations
- AutoLink: Self-supervised Learning of Human Skeletons and Object Outlines by Linking KeypointsXingzhe He, Bastian Wandt, Helge RhodinNeurIPS 2022 · 28 citations
- Weakly-supervised 3D Pose Transfer with KeypointsJinnan Chen, Chen Li, Gim Hee LeeICCV 2023 · 13 citations
- Unsupervised 3D Structure Inference from Category-Specific Image CollectionsWeikang Wang, Dongliang Cao, Florian BernardCVPR 2024
