Model-Aware Gesture-to-Gesture Translation
Hezhen Hu, Weilun Wang, Wengang Zhou, Weichao Zhao, Houqiang Li
摘要
Hand gesture-to-gesture translation is a significant and interesting problem, which serves as a key role in many applications, such as sign language production. This task involves fine-grained structure understanding of the mapping between the source and target gestures. Current works follow a data-driven paradigm based on sparse 2D joint representation. However, given the insufficient representation capability of 2D joints, this paradigm easily leads to blurry generation results with incorrect structure. In this paper, we propose a novel model-aware gesture-to-gesture translation framework, which introduces hand prior with hand meshes as the intermediate representation. To take full advantage of the structured hand model, we first build a dense topology map aligning the image plane with the encoded embedding of the visible hand mesh. Then, a transformation flow is calculated based on the correspondence of the source and target topology map. During the generation stage, we inject the topology information into generation streams by modulating the activations in a spatially-adaptive manner. Further, we incorporate the source local characteristic to enhance the translated gesture image according to the transformation flow. Extensive experiments on two benchmark datasets have demonstrated that our method achieves new state-of-the-art performance.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了最后一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper3
- SignBERT: Pre-Training of Hand-Model-Aware Representation for Sign Language RecognitionHezhen Hu, Weichao Zhao, Wengang Zhou, Yuechen Wang 等ICCV 2021 · 被引用 125 次
- Hand-Object Interaction Image GenerationHezhen Hu, Weilun Wang, Wengang Zhou, Houqiang LiNeurIPS 2022 · 被引用 24 次
- FoundHand: Large-Scale Domain-Specific Learning for Controllable Hand Image GenerationKefan Chen, Chaerin Min, Linguang Zhang, Shreyas Hampali 等CVPR 2025
它引用的顶会 Paper6
- Liquid Warping GAN: A Unified Framework for Human Motion Imitation, Appearance Transfer and Novel View SynthesisWen Liu, Zhixin Piao, Jie Min, Wenhan Luo 等ICCV 2019 · 被引用 285 次
- SO-HandNet: Self-Organizing Network for 3D Hand Pose Estimation With Semi-Supervised LearningYujin Chen, Zhigang Tu, Liuhao Ge, Dejun Zhang 等ICCV 2019 · 被引用 87 次
- MM-Hand: 3D-Aware Multi-Modal Guided Hand Generation for 3D Hand Pose SynthesisZhenyu Wu, Duc Hoang, Shih-Yao Lin, Yusheng Xie 等ACM MM 2020 · 被引用 16 次
- DeepCap: Monocular Human Performance Capture Using Weak SupervisionMarc Habermann, Weipeng Xu, Michael Zollhöfer, Gerard Pons-Moll 等CVPR 2020
- Disentangled and Controllable Face Image Generation via 3D Imitative-Contrastive LearningYu Deng, Jiaolong Yang, Dong Chen, Fang Wen 等CVPR 2020
相关 Paper
- Hand-Model-Aware Sign Language RecognitionHezhen Hu, Wengang Zhou, Houqiang LiAAAI 2021 · 被引用 79 次
- LLM Knows Body Language, Too: Translating Speech Voices into Human GesturesChenghao Xu, Guangtao Lyu, Jiexi Yan, Muli Yang 等ACL 2024
- Robust Photo-Realistic Hand Gesture Generation: from Single View to Multiple ViewQifan Fu, Xu Chen, Muhammad Asad, Shanxin Yuan 等ACM MM 2025
- Structure-aware Person Image Generation with Pose Decomposition and Semantic CorrelationJilin Tang, Yi Yuan, Tianjia Shao, Yong Liu 等AAAI 2021 · 被引用 22 次
- SignPR: A Progressive Vector-Quantized Diffusion Framework for Sign Language ProductionXiao Liu, Shiwei Gan, Yafeng Yin, Bowen Guo 等CVPR 2026 · 被引用 2 次
