Cross-Domain 3D Hand Pose Estimation with Dual Modalities
Qiuxia Lin, Linlin Yang, Angela Yao
摘要
Recent advances in hand pose estimation have shed light on utilizing synthetic data to train neural networks, which however inevitably hinders generalization to realworld data due to domain gaps. To solve this problem, we present a framework for cross-domain semi-supervised hand pose estimation and target the challenging scenario of learning models from labelled multi-modal synthetic data and unlabelled real-world data. To that end, we propose a dual-modality network that exploits synthetic RGB and synthetic depth images. For pre-training, our network uses multi-modal contrastive learning and attention-fused supervision to learn effective representations of the RGB images. We then integrate a novel self-distillation technique during fine-tuning to reduce pseudo-label noise. Experiments show that the proposed method significantly improves 3D hand pose estimation and 2D keypoint detection on benchmarks.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了最后一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper8
- Triangulation Residual Loss for Data-efficient 3D Pose EstimationJiachen Zhao, Tao Yu, Liang An, Yipeng Huang 等NeurIPS 2023 · 被引用 13 次
- Touchscreen-based Hand Tracking for Remote Whiteboard InteractionXinshuang Liu, Yizhong Zhang, Xin TongUIST 2024 · 被引用 8 次
- SiMA-Hand: Boosting 3D Hand-Mesh Reconstruction by Single-to-Multi-View AdaptationYinqiao Wang, Hao Xu, Pheng-Ann Heng, Chi-Wing FuAAAI 2024 · 被引用 5 次
- Synthetic-to-Real Pose Estimation with Geometric ReconstructionQiuxia Lin, Kerui Gu, Linlin Yang, Angela YaoNeurIPS 2023 · 被引用 4 次
- Analyzing the Synthetic-to-Real Domain Gap in 3D Hand Pose EstimationZhuoran Zhao, Linlin Yang, Pengzhan Sun, Pan Hui 等CVPR 2025
它引用的顶会 Paper18
- A Simple Framework for Contrastive Learning of Visual RepresentationsTing Chen, Simon Kornblith, Mohammad Norouzi, Geoffrey E. HintonICML 2020 · 被引用 24,064 次
- Supervised Contrastive LearningPrannay Khosla, Piotr Teterwak, Chen Wang, Aaron Sarna 等NeurIPS 2020 · 被引用 7,049 次
- Understanding Contrastive Representation Learning through Alignment and Uniformity on the HypersphereTongzhou Wang, Phillip IsolaICML 2020 · 被引用 2,360 次
- FreiHAND: A Dataset for Markerless Capture of Hand Pose and Shape From Single RGB ImagesChristian Zimmermann, Duygu Ceylan, Jimei Yang, Bryan C. Russell 等ICCV 2019 · 被引用 493 次
- Exploring Balanced Feature Spaces for Representation LearningBingyi Kang, Yu Li, Sa Xie, Zehuan Yuan 等ICLR 2021 · 被引用 296 次
相关 Paper
- SemiHand: Semi-supervised Hand Pose Estimation with ConsistencyLinlin Yang, Shicheng Chen, Angela YaoICCV 2021 · 被引用 42 次
- Knowledge As Priors: Cross-Modal Knowledge Generalization for Datasets Without Superior KnowledgeLong Zhao, Xi Peng, Yuxiao Chen, Mubbasir Kapadia 等CVPR 2020
- Cross-Domain and Cross-Modal Knowledge Distillation in Domain Adaptation for 3D Semantic SegmentationMiaoyu Li, Yachao Zhang, Yuan Xie, Zuodong Gao 等ACM MM 2022 · 被引用 30 次
- Rule Meets Learning: Confidence-Aware Multi-View Fusion for Self-Supervised 3D Hand Pose EstimationPengfei Ren, Jingyu Wang, Haifeng Sun, Qi Qi 等ACM MM 2025 · 被引用 1 次
- Keypoint Fusion for RGB-D Based 3D Hand Pose EstimationXingyu Liu, Pengfei Ren, Yuanyuan Gao, Jingyu Wang 等AAAI 2024 · 被引用 11 次
