TransMoMo: Invariance-Driven Unsupervised Video Motion Retargeting
Zhuoqian Yang, Wentao Zhu, Wayne Wu, Chen Qian, Qiang Zhou, Bolei Zhou, Chen Change Loy
摘要
Abstract We present a lightweight video motion retargeting approach TransMoMo that is capable of transferring motion of a person in a source video realistically to another video of a target person (Fig. 1 ). Without using any paired data for supervision, the proposed method can be trained in an unsupervised manner by exploiting invariance properties of three orthogonal factors of variation including motion, structure, and view-angle. Specifically, with loss functions carefully derived based on invariance, we train an autoencoder to disentangle the latent representations of such factors given the source and target video clips. This allows us to selectively transfer motion extracted from the source video seamlessly to the target video in spite of structural and view-angle disparities between the source and the target. The relaxed assumption of paired data allows our method to be trained on a vast amount of videos needless of manual annotation of source-target pairing, leading to improved robustness against large structural variations and extreme motion in videos. We demonstrate the effectiveness of our method over the state-of-the-art methods such as NKN [41] , EDN [8] and LCM [4] . Code, model and data are publicly available on our project page. 1
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper20
- MotionBERT: A Unified Perspective on Learning Human Motion RepresentationsWentao Zhu, Xiaoxuan Ma, Zhaoyang Liu, Libin Liu 等ICCV 2023 · 被引用 322 次
- Latent Image Animator: Learning to Animate Images via Latent Space NavigationYaohui Wang, Di Yang, François Brémond, Antitza DantchevaICLR 2022 · 被引用 219 次
- Zero-shot Synthesis with Group-Supervised LearningYunhao Ge, Sami Abu-El-Haija, Gan Xin, Laurent IttiICLR 2021 · 被引用 45 次
- FashionMirror: Co-attention Feature-remapping Virtual Try-on with Sequential Template PosesChieh-Yun Chen, Ling Lo, Pin-Jui Huang, Hong-Han Shuai 等ICCV 2021 · 被引用 34 次
- Structure-Aware Motion Transfer with Deformable Anchor ModelJiale Tao, Biao Wang, Borun Xu, Tiezheng Ge 等CVPR 2022 · 被引用 33 次
它引用的顶会 Paper3
- Everybody Dance NowCaroline Chan, Shiry Ginosar, Tinghui Zhou, Alexei A. EfrosICCV 2019 · 被引用 840 次
- Liquid Warping GAN: A Unified Framework for Human Motion Imitation, Appearance Transfer and Novel View SynthesisWen Liu, Zhixin Piao, Jie Min, Wenhan Luo 等ICCV 2019 · 被引用 285 次
- Make a Face: Towards Arbitrary High Fidelity Face ManipulationShengju Qian, Kwan-Yee Lin, Wayne Wu, Yangxiaokang Liu 等ICCV 2019 · 被引用 75 次
相关 Paper
- MoCaNet: Motion Retargeting In-the-Wild via Canonicalization NetworksWentao Zhu, Zhuoqian Yang, Ziang Di, Wayne Wu 等AAAI 2022 · 被引用 24 次
- Flow Guided Transformable Bottleneck Networks for Motion RetargetingJian Ren, Menglei Chai, Oliver J. Woodford, Kyle Olszewski 等CVPR 2021
- MotionShot: Adaptive Motion Transfer Across Arbitrary Objects for Text-to-Video GenerationYanchen Liu, Yanan Sun, Zhening Xing, Junyao Gao 等ICCV 2025 · 被引用 5 次
- Unpaired motion style transfer from video to animationKfir Aberman, Yijia Weng, Dani Lischinski, Daniel Cohen-Or 等SIGGRAPH 2020 · 被引用 178 次
- FSRT: Facial Scene Representation Transformer for Face Reenactment from Factorized Appearance, Head-Pose, and Facial Expression FeaturesAndre Rochow, Max Schwarz, Sven BehnkeCVPR 2024 · 被引用 17 次
