Self-Supervised 3D Hand Pose Estimation from monocular RGB via Contrastive Learning
Adrian Spurr, Aneesh Dahiya, Xi Wang, Xucong Zhang, Otmar Hilliges
摘要
Encouraged by the success of contrastive learning on image classification tasks, we propose a new self-supervised method for the structured regression task of 3D hand pose estimation. Contrastive learning makes use of unlabeled data for the purpose of representation learning via a loss formulation that encourages the learned feature representations to be invariant under any image transformation. For 3D hand pose estimation, it too is desirable to have invariance to appearance transformation such as color jitter. However, the task requires equivariance under affine transformations, such as rotation and translation. To address this issue, we propose an equivariant contrastive objective and demonstrate its effectiveness in the context of 3D hand pose estimation. We experimentally investigate the impact of invariant and equivariant contrastive objectives and show that learning equivariant features leads to better representations for the task of 3D hand pose estimation. Furthermore, we show that standard ResNets with sufficient depth, trained on additional unlabeled data, attain improvements of up to 14.5% in PA-EPE on FreiHAND and thus achieves state-of-the-art performance without any task specific, specialized architectures. Code and models are available at https://ait.ethz.ch/projects/2021/PeCLR/
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper20
- MobRecon: Mobile-Friendly Hand Mesh Reconstruction from Monocular ImageXingyu Chen, Yufeng Liu, Yajiao Dong, Xiong Zhang 等CVPR 2022 · 被引用 97 次
- Contrastive Regression for Domain Adaptation on Gaze EstimationYaoming Wang, Yangzhou Jiang, Jin Li, Bingbing Ni 等CVPR 2022 · 被引用 80 次
- Towards Robust and Expressive Whole-body Human Pose and Shape EstimationHui En Pang, Zhongang Cai, Lei Yang, Qingyi Tao 等NeurIPS 2023 · 被引用 16 次
- Depth-discriminative Metric Learning for Monocular 3D Object DetectionWonhyeok Choi, Mingyu Shin, Sunghoon ImNeurIPS 2023 · 被引用 13 次
- OCR-Pose: Occlusion-aware Contrastive Representation for Unsupervised 3D Human Pose EstimationJunjie Wang, Zhenbo Yu, Zhengyan Tong, Hang Wang 等ACM MM 2022 · 被引用 12 次
它引用的顶会 Paper10
- A Simple Framework for Contrastive Learning of Visual RepresentationsTing Chen, Simon Kornblith, Mohammad Norouzi, Geoffrey E. HintonICML 2020 · 被引用 24,064 次
- Big Self-Supervised Models are Strong Semi-Supervised LearnersTing Chen, Simon Kornblith, Kevin Swersky, Mohammad Norouzi 等NeurIPS 2020 · 被引用 2,611 次
- Data-Efficient Image Recognition with Contrastive Predictive CodingOlivier J. HénaffICML 2020 · 被引用 1,553 次
- FreiHAND: A Dataset for Markerless Capture of Hand Pose and Shape From Single RGB ImagesChristian Zimmermann, Duygu Ceylan, Jimei Yang, Bryan C. Russell 等ICCV 2019 · 被引用 493 次
- End-to-End Hand Mesh Recovery From a Monocular RGB ImageXiong Zhang, Qiang Li, Hong Mo, Wenbo Zhang 等ICCV 2019 · 被引用 248 次
相关 Paper
- SiMHand: Mining Similar Hands for Large-Scale 3D Hand Pose Pre-trainingNie Lin, Takehiko Ohkawa, Yifei Huang, Mingfang Zhang 等ICLR 2025
- CycleHand: Increasing 3D Pose Estimation Ability on In-the-wild Monocular Image through Cyclic FlowDaiheng Gao, Xindi Zhang, Xingyu Chen, Andong Tan 等ACM MM 2022 · 被引用 5 次
- On Equivariant and Invariant Learning of Object Landmark RepresentationsZezhou Cheng, Jong-Chyi Su, Subhransu MajiICCV 2021 · 被引用 18 次
- Cross-Domain 3D Hand Pose Estimation with Dual ModalitiesQiuxia Lin, Linlin Yang, Angela YaoCVPR 2023
- CrossPoint: Self-Supervised Cross-Modal Contrastive Learning for 3D Point Cloud UnderstandingMohamed Afham, Isuru Dissanayake, Dinithi Dissanayake, Amaya Dharmasiri 等CVPR 2022 · 被引用 286 次
