CycleHand: Increasing 3D Pose Estimation Ability on In-the-wild Monocular Image through Cyclic Flow
Daiheng Gao, Xindi Zhang, Xingyu Chen, Andong Tan, Bang Zhang, Pan Pan, Ping Tan
摘要
Current methods for 3D hand pose estimation fail to generalize well to in-the-wild new scenarios due to varying camera viewpoints, self-occlusions, and complex environments. To address this problem, we propose CycleHand to improve the generalization ability of the model in a self-supervised manner. Our motivation is based on an observation: if one globally rotates the whole hand and reversely rotates it back, the estimated 3D poses of fingers should keep consistent before and after the rotation because the wrist-relative hand poses stay unchanged during global 3D rotation. Hence, we propose arbitrary-rotation self-supervised consistency learning to improve the model's robustness for varying viewpoints. Another innovation of CycleHand is that we propose a high-fidelity texture map to render the photorealistic rotated hand with different lighting conditions, backgrounds, and skin tones to further enhance the effectiveness of our self-supervised task. To reduce the potential negative effects brought by the domain shift of synthetic images, we use the idea of contrastive learning to learn a synthetic-real consistent feature extractor in extracting domain-irrelevant hand representations. Experiments show that CycleHand can largely improve the hand pose estimation performance in both canonical datasets and real-world applications.
问问这篇 Paper
问问你的智能体。
Lune 读过与它相关的顶会 Paper,每个回答都会注明依据哪几篇。
引用它的顶会 Paper4
- CLIP-Hand3D: Exploiting 3D Hand Pose Estimation via Context-Aware PromptingShaoxiang Guo, Qing Cai, Lin Qi, Junyu DongACM MM 2023 · 被引用 10 次
- OHTA: One-shot Hand Avatar via Data-driven Implicit PriorsXiaozheng Zheng, Chao Wen, Zhuo Su, Zeran Xu 等CVPR 2024 · 被引用 7 次
- Diffusion-Based 3D Hand Motion Recovery with Intuitive PhysicsYufei Zhang, Zijun Cui, Jeffrey O. Kephart, Qiang JiICCV 2025 · 被引用 1 次
- Hand Avatar: Free-Pose Hand Animation and Rendering from Monocular VideoXingyu Chen, Baoyuan Wang, Heung-Yeung ShumCVPR 2023
相关 Paper
- Self-Supervised 3D Hand Pose Estimation from monocular RGB via Contrastive LearningAdrian Spurr, Aneesh Dahiya, Xi Wang, Xucong Zhang 等ICCV 2021 · 被引用 74 次
- Model-Based 3D Hand Reconstruction via Self-Supervised LearningYujin Chen, Zhigang Tu, Di Kang, Linchao Bao 等CVPR 2021
- SiMHand: Mining Similar Hands for Large-Scale 3D Hand Pose Pre-trainingNie Lin, Takehiko Ohkawa, Yifei Huang, Mingfang Zhang 等ICLR 2025
- SemiHand: Semi-supervised Hand Pose Estimation with ConsistencyLinlin Yang, Shicheng Chen, Angela YaoICCV 2021 · 被引用 42 次
- Cross-Domain 3D Hand Pose Estimation with Dual ModalitiesQiuxia Lin, Linlin Yang, Angela YaoCVPR 2023
