Semantics-Aware Motion Retargeting with Vision-Language Models
Haodong Zhang, Zhike Chen, Haocheng Xu, Lei Hao, Xiaofei Wu, Songcen Xu, Zhensong Zhang, Yue Wang, Rong Xiong
摘要
Capturing and preserving motion semantics is essential to motion retargeting between animation characters. However, most of the previous works neglect the semantic information or rely on human-designed joint-level representations. Here, we present a novel Semantics-aware Motion reTargeting (SMT) method with the advantage of vision-language models to extract and maintain meaningful motion semantics. We utilize a differentiable module to ren-der 3D motions. Then the high-level motion semantics are incorporated into the motion retargeting process by feeding the vision-language model with the rendered images and aligning the extracted semantic embeddings. To en-sure the preservation of fine-grained motion details and high-level semantics, we adopt a two-stage pipeline consisting of skeleton-aware pretraining and fine-tuning with semantics and geometry constraints. Experimental results show the effectiveness of the proposed method in producing high-quality motion retargeting results while accurately preserving motion semantics. Project page can be found at https://sites.google.com/view/smtnet.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper4
- Skinned Motion Retargeting with Dense Geometric Interaction PerceptionZijie Ye, Jia-Wei Liu, Jia Jia, Shikun Sun 等NeurIPS 2024 · 被引用 22 次
- STaR: Seamless Spatial-Temporal Aware Motion Retargeting with Penetration and Consistency ConstraintsXiaohang Yang, Qing Wang, Jiahao Yang, Gregory G. Slabaugh 等ICCV 2025 · 被引用 2 次
- Text-to-Any-Skeleton Motion Generation Without RetargetingQingyuan Liu, Ke Lu, Kun Dong, Jian Xue 等ICCV 2025 · 被引用 1 次
- AniMimic: Imitating 3D Animation from Video PriorsTianyi Xie, Yunuo Chen, Yaowei Guo, Yin Yang 等CVPR 2026
它引用的顶会 Paper13
- Learning Transferable Visual Models From Natural Language SupervisionAlec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh 等ICML 2021 · 被引用 47,906 次
- BLIP-2: Bootstrapping Language-Image Pre-training with Frozen Image Encoders and Large Language ModelsJunnan Li, Dongxu Li, Silvio Savarese, Steven C. H. HoiICML 2023 · 被引用 7,873 次
- MiniGPT-4: Enhancing Vision-Language Understanding with Advanced Large Language ModelsDeyao Zhu, Jun Chen, Xiaoqian Shen, Xiang Li 等ICLR 2024 · 被引用 3,079 次
- Soft Rasterizer: A Differentiable Renderer for Image-Based 3D ReasoningShichen Liu, Weikai Chen, Tianye Li, Hao LiICCV 2019 · 被引用 789 次
- Generating Diverse and Natural 3D Human Motions from TextChuan Guo, Shihao Zou, Xinxin Zuo, Sen Wang 等CVPR 2022 · 被引用 462 次
相关 Paper
- Skinned Motion Retargeting with Spatially Adaptive Interaction GuidanceSoojin Choi, Seokhyeon Hong, Chaelin Kim, Junghyun Nam 等SIGGRAPH 2026
- Skinned Motion Retargeting with Residual Perception of Motion Semantics & GeometryJiaxu Zhang, Junwu Weng, Di Kang, Fang Zhao 等CVPR 2023
- Semantic-Aware Motion Encoding for Topology-Agnostic Character AnimationZongye Zhang, Yuzhuo Cui, Qingjie Liu, Yunhong WangICML 2026 · 被引用 1 次
- Motion-Aligned Word Embeddings for Text-to-Motion GenerationKe Han, Yueming Lyu, Nicu SebeICLR 2026
- Contact-Aware Retargeting of Skinned MotionRuben Villegas, Duygu Ceylan, Aaron Hertzmann, Jimei Yang 等ICCV 2021 · 被引用 52 次
