Model-Aware Gesture-to-Gesture Translation
Hezhen Hu, Weilun Wang, Wengang Zhou, Weichao Zhao, Houqiang Li
Abstract
Hand gesture-to-gesture translation is a significant and interesting problem, which serves as a key role in many applications, such as sign language production. This task involves fine-grained structure understanding of the mapping between the source and target gestures. Current works follow a data-driven paradigm based on sparse 2D joint representation. However, given the insufficient representation capability of 2D joints, this paradigm easily leads to blurry generation results with incorrect structure. In this paper, we propose a novel model-aware gesture-to-gesture translation framework, which introduces hand prior with hand meshes as the intermediate representation. To take full advantage of the structured hand model, we first build a dense topology map aligning the image plane with the encoded embedding of the visible hand mesh. Then, a transformation flow is calculated based on the correspondence of the source and target topology map. During the generation stage, we inject the topology information into generation streams by modulating the activations in a spatially-adaptive manner. Further, we incorporate the source local characteristic to enhance the translated gesture image according to the transformation flow. Extensive experiments on two benchmark datasets have demonstrated that our method achieves new state-of-the-art performance.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 03994c03-767e-46cb-8f36-e02f65e1e4efCited by top-tier papers3
- SignBERT: Pre-Training of Hand-Model-Aware Representation for Sign Language RecognitionHezhen Hu, Weichao Zhao, Wengang Zhou, Yuechen Wang et al.ICCV 2021 · 125 citations
- Hand-Object Interaction Image GenerationHezhen Hu, Weilun Wang, Wengang Zhou, Houqiang LiNeurIPS 2022 · 24 citations
- FoundHand: Large-Scale Domain-Specific Learning for Controllable Hand Image GenerationKefan Chen, Chaerin Min, Linguang Zhang, Shreyas Hampali et al.CVPR 2025
Builds on6
- Liquid Warping GAN: A Unified Framework for Human Motion Imitation, Appearance Transfer and Novel View SynthesisWen Liu, Zhixin Piao, Jie Min, Wenhan Luo et al.ICCV 2019 · 285 citations
- SO-HandNet: Self-Organizing Network for 3D Hand Pose Estimation With Semi-Supervised LearningYujin Chen, Zhigang Tu, Liuhao Ge, Dejun Zhang et al.ICCV 2019 · 87 citations
- MM-Hand: 3D-Aware Multi-Modal Guided Hand Generation for 3D Hand Pose SynthesisZhenyu Wu, Duc Hoang, Shih-Yao Lin, Yusheng Xie et al.ACM MM 2020 · 16 citations
- DeepCap: Monocular Human Performance Capture Using Weak SupervisionMarc Habermann, Weipeng Xu, Michael Zollhöfer, Gerard Pons-Moll et al.CVPR 2020
- Disentangled and Controllable Face Image Generation via 3D Imitative-Contrastive LearningYu Deng, Jiaolong Yang, Dong Chen, Fang Wen et al.CVPR 2020
Related papers
- Hand-Model-Aware Sign Language RecognitionHezhen Hu, Wengang Zhou, Houqiang LiAAAI 2021 · 79 citations
- LLM Knows Body Language, Too: Translating Speech Voices into Human GesturesChenghao Xu, Guangtao Lyu, Jiexi Yan, Muli Yang et al.ACL 2024
- Robust Photo-Realistic Hand Gesture Generation: from Single View to Multiple ViewQifan Fu, Xu Chen, Muhammad Asad, Shanxin Yuan et al.ACM MM 2025
- Structure-aware Person Image Generation with Pose Decomposition and Semantic CorrelationJilin Tang, Yi Yuan, Tianjia Shao, Yong Liu et al.AAAI 2021 · 22 citations
- SignPR: A Progressive Vector-Quantized Diffusion Framework for Sign Language ProductionXiao Liu, Shiwei Gan, Yafeng Yin, Bowen Guo et al.CVPR 2026 · 2 citations
