MM-Hand: 3D-Aware Multi-Modal Guided Hand Generation for 3D Hand Pose Synthesis
Zhenyu Wu, Duc Hoang, Shih-Yao Lin, Yusheng Xie, Liangjian Chen, Yen-Yu Lin, Zhangyang Wang, Wei Fan
Abstract
Estimating the 3D hand pose from a monocular RGB image is important but challenging. A solution is training on large-scale RGB hand images with accurate 3D hand keypoint annotations. However, it is too expensive in practice. Instead, we have developed a learning-based approach to synthesize realistic, diverse, and 3D pose-preserving hand images under the guidance of 3D pose information. We propose a 3D-aware multi-modal guided hand generative network (MM-Hand), together with a novel geometrybased curriculum learning strategy. Our extensive experimental results demonstrate that the 3D-annotated images generated by MM-Hand qualitatively and quantitatively outperform existing options. Moreover, the augmented data can consistently improve the quantitative performance of the state-of-the-art 3D hand pose estimators on two benchmark datasets. The code will be available at https://github.com/ScottHoang/mm-hand.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 237244ea-0d3d-481f-8cea-68a7c1faee85Cited by top-tier papers4
- Hand-Object Interaction Image GenerationHezhen Hu, Weilun Wang, Wengang Zhou, Houqiang LiNeurIPS 2022 · 24 citations
- CLIP-Hand3D: Exploiting 3D Hand Pose Estimation via Context-Aware PromptingShaoxiang Guo, Qing Cai, Lin Qi, Junyu DongACM MM 2023 · 10 citations
- Single Image 3D Object Estimation with Primitive Graph NetworksQian He, Desen Zhou, Bo Wan, Xuming HeACM MM 2021 · 1 citation
- Model-Aware Gesture-to-Gesture TranslationHezhen Hu, Weilun Wang, Wengang Zhou, Weichao Zhao et al.CVPR 2021
Builds on8
- DeblurGAN-v2: Deblurring (Orders-of-Magnitude) Faster and BetterOrest Kupyn, Tetiana Martyniuk, Junru Wu, Zhangyang WangICCV 2019 · 1,100 citations
- Everybody Dance NowCaroline Chan, Shiry Ginosar, Tinghui Zhou, Alexei A. EfrosICCV 2019 · 840 citations
- FreiHAND: A Dataset for Markerless Capture of Hand Pose and Shape From Single RGB ImagesChristian Zimmermann, Duygu Ceylan, Jimei Yang, Bryan C. Russell et al.ICCV 2019 · 493 citations
- Liquid Warping GAN: A Unified Framework for Human Motion Imitation, Appearance Transfer and Novel View SynthesisWen Liu, Zhixin Piao, Jie Min, Wenhan Luo et al.ICCV 2019 · 285 citations
- Controllable Artistic Text Style Transfer via Shape-Matching GANShuai Yang, Zhangyang Wang, Zhaowen Wang, Ning Xu et al.ICCV 2019 · 110 citations
Related papers
- Model-Based 3D Hand Reconstruction via Self-Supervised LearningYujin Chen, Zhigang Tu, Di Kang, Linchao Bao et al.CVPR 2021
- SemiHand: Semi-supervised Hand Pose Estimation with ConsistencyLinlin Yang, Shicheng Chen, Angela YaoICCV 2021 · 42 citations
- End-to-End Hand Mesh Recovery From a Monocular RGB ImageXiong Zhang, Qiang Li, Hong Mo, Wenbo Zhang et al.ICCV 2019 · 248 citations
- HOnnotate: A Method for 3D Annotation of Hand and Object PosesShreyas Hampali, Mahdi Rad, Markus Oberweger, Vincent LepetitCVPR 2020
- RenderIH: A Large-scale Synthetic Dataset for 3D Interacting Hand Pose EstimationLijun Li, Linrui Tian, Xindi Zhang, Qi Wang et al.ICCV 2023 · 28 citations
