Two Heads Are Better than One: Image-Point Cloud Network for Depth-Based 3D Hand Pose Estimation
Pengfei Ren, Yuchen Chen, Jiachang Hao, Haifeng Sun, Qi Qi, Jingyu Wang, Jianxin Liao
摘要
Depth images and point clouds are the two most commonly used data representations for depth-based 3D hand pose estimation. Benefiting from the structuring of image data and the inherent inductive biases of the 2D Convolutional Neural Network (CNN), image-based methods are highly efficient and effective. However, treating the depth data as a 2D image inevitably ignores the 3D nature of depth data. Point cloud-based methods can better mine the 3D geometric structure of depth data. However, these methods suffer from the disorder and non-structure of point cloud data, which is computationally inefficient. In this paper, we propose an Image-Point cloud Network (IPNet) for accurate and robust 3D hand pose estimation. IPNet utilizes 2D CNN to extract visual representations in 2D image space and performs iterative correction in 3D point cloud space to exploit the 3D geometry information of depth data. In particular, we propose a sparse anchor-based "aggregation-interaction-propagation'' paradigm to enhance point cloud features and refine the hand pose, which reduces irregular data access. Furthermore, we introduce a 3D hand model to the iterative correction process, which significantly improves the robustness of IPNet to occlusion and depth holes. Experiments show that IPNet outperforms state-of-the-art methods on three challenging hand datasets.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper11
- Keypoint Fusion for RGB-D Based 3D Hand Pose EstimationXingyu Liu, Pengfei Ren, Yuanyuan Gao, Jingyu Wang 等AAAI 2024 · 被引用 11 次
- Hierarchical-Aware Orthogonal Disentanglement Framework for Fine-Grained Skeleton-Based Action RecognitionHaochen Chang, Pengfei Ren, Haoyang Zhang, Liang Xie 等ICCV 2025 · 被引用 8 次
- Fine-Grained Multi-View Hand Reconstruction Using Inverse RenderingQijun Gan, Wentong Li, Jinwei Ren, Jianke ZhuAAAI 2024 · 被引用 7 次
- EchoDiffusion: Waveform Conditioned Diffusion Models for Echo-Based Depth EstimationWenjie Zhang, Jun Yin, Long Ma, Peng Yu 等AAAI 2025 · 被引用 2 次
- Monocular 3D Hand Mesh Recovery via Dual Noise EstimationHanhui Li, Xiaojian Lin, Xuan Huang, Zejun Yang 等AAAI 2024 · 被引用 2 次
它引用的顶会 Paper19
- A ConvNet for the 2020sZhuang Liu, Hanzi Mao, Chao-Yuan Wu, Christoph Feichtenhofer 等CVPR 2022 · 被引用 6,782 次
- KPConv: Flexible and Deformable Convolution for Point CloudsHugues Thomas, Charles R. Qi, Jean-Emmanuel Deschaud, Beatriz Marcotegui 等ICCV 2019 · 被引用 3,193 次
- Rethinking Network Design and Local Geometry in Point Cloud: A Simple Residual MLP FrameworkXu Ma, Can Qin, Haoxuan You, Haoxi Ran 等ICLR 2022 · 被引用 841 次
- PyMAF: 3D Human Pose and Shape Regression with Pyramidal Mesh Alignment Feedback LoopHongwen Zhang, Yating Tian, Xinchi Zhou, Wanli Ouyang 等ICCV 2021 · 被引用 376 次
- Surface Representation for Point CloudsHaoxi Ran, Jun Liu, Chengjie WangCVPR 2022 · 被引用 230 次
相关 Paper
- SO-HandNet: Self-Organizing Network for 3D Hand Pose Estimation With Semi-Supervised LearningYujin Chen, Zhigang Tu, Liuhao Ge, Dejun Zhang 等ICCV 2019 · 被引用 87 次
- HandFoldingNet: A 3D Hand Pose Estimation Network Using Multiscale-Feature Guided Folding of a 2D Hand SkeletonWencan Cheng, Jae Hyun Park, Jong Hwan KoICCV 2021 · 被引用 50 次
- HandVoxNet: Deep Voxel-Based Network for 3D Hand Shape and Pose Estimation From a Single Depth MapJameel Malik, Ibrahim Abdelaziz, Ahmed Elhayek, Soshi Shimada 等CVPR 2020
- Weakly Supervised Adversarial Learning for 3D Human Pose Estimation from Point CloudsZihao Zhang, Lei Hu, Xiaoming Deng, Shihong XiaIEEE VR 2020 · 被引用 51 次
- A2J: Anchor-to-Joint Regression Network for 3D Articulated Pose Estimation From a Single Depth ImageFu Xiong, Boshen Zhang, Yang Xiao, Zhiguo Cao 等ICCV 2019 · 被引用 178 次
