Few-Shot Geometry-Aware Keypoint Localization
Xingzhe He, Gaurav Bharaj, David Ferman, Helge Rhodin, Pablo Garrido
Abstract
Supervised keypoint localization methods rely on large manually labeled image datasets, where objects can deform, articulate, or occlude. However, creating such large keypoint labels is time-consuming and costly, and is often error-prone due to inconsistent labeling. Thus, we desire an approach that can learn keypoint localization with fewer yet consistently annotated images. To this end, we present a novel formulation that learns to localize semantically consistent keypoint definitions, even for occluded regions, for varying object categories. We use a few user-labeled 2D images as input examples, which are extended via self-supervision using a larger unlabeled dataset. Unlike unsupervised methods, the few-shot images act as semantic shape constraints for object localization. Furthermore, we introduce 3D geometryaware constraints to uplift keypoints, achieving more accurate 2D localization. Our general-purpose formulation paves the way for semantically conditioned generative modeling and attains competitive or state-of-the-art accuracy on several datasets, including human faces, eyes, animals, cars, and never-before-seen mouth interior (teeth) localization tasks, not attempted by the previous few-shot methods.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 3ca5be37-0e58-43e7-a9e4-6c574651e952Cited by top-tier papers6
- Detect Any Keypoints: An Efficient Light-Weight Few-Shot Keypoint DetectorChangsheng Lu, Piotr KoniuszAAAI 2024 · 12 citations
- Weak-shot Keypoint Estimation via Keyness and Correspondence TransferJunjie Chen, Zeyu Luo, Zezheng Liu, Wenhui Jiang et al.NeurIPS 2025 · 5 citations
- Unsupervised 3D Structure Inference from Category-Specific Image CollectionsWeikang Wang, Dongliang Cao, Florian BernardCVPR 2024
- Doodle Your Keypoints: Sketch-Based Few-Shot Keypoint DetectionSubhajit Maity, Ayan Kumar Bhunia, Subhadeep Koley, Pinaki Nath Chowdhury et al.ICCV 2025
- Incremental Object Keypoint LearningMingfu Liang, Jiahuan Zhou, Xu Zou, Ying WuCVPR 2025
Builds on28
- Denoising Diffusion Probabilistic ModelsJonathan Ho, Ajay Jain, Pieter AbbeelNeurIPS 2020 · 35,902 citations
- Fake it till you make it: face analysis in the wild using synthetic data aloneErroll Wood, Tadas Baltrusaitis, Charlie Hewitt, Sebastian Dziadzio et al.ICCV 2021 · 331 citations
- Human Pose Regression with Residual Log-likelihood EstimationJiefeng Li, Siyuan Bian, Ailing Zeng, Can Wang et al.ICCV 2021 · 286 citations
- StyleSDF: High-Resolution 3D-Consistent Image and Geometry GenerationRoy Or-El, Xuan Luo, Mengyi Shan, Eli Shechtman et al.CVPR 2022 · 229 citations
- Splicing ViT Features for Semantic Appearance TransferNarek Tumanyan, Omer Bar-Tal, Shai Bagon, Tali DekelCVPR 2022 · 128 citations
Related papers
- Semi-supervised Keypoint LocalizationOlga Moskvyak, Frédéric Maire, Feras Dayoub, Mahsa BaktashmotlaghICLR 2021 · 17 citations
- Generalizable Object Keypoint Localization from Generative PriorsDongkai Wang, Jiang Duan, Liangjian Wen, Shiyu Xuan et al.CVPR 2025
- Pseudo-Labeled Auto-Curriculum Learning for Semi-Supervised Keypoint LocalizationCan Wang, Sheng Jin, Yingda Guan, Wentao Liu et al.ICLR 2022 · 17 citations
- Back to 3D: Few-Shot 3D Keypoint Detection with Back-Projected 2D FeaturesThomas Wimmer, Peter Wonka, Maks OvsjanikovCVPR 2024
- Self-Supervised Image Representation Learning with Geometric Set ConsistencyNenglun Chen, Lei Chu, Hao Pan, Yan Lu et al.CVPR 2022 · 8 citations
