Structure-Aware Motion Transfer with Deformable Anchor Model
Jiale Tao, Biao Wang, Borun Xu, Tiezheng Ge, Yuning Jiang, Wen Li, Lixin Duan
摘要
Given a source image and a driving video depicting the same object type, the motion transfer task aims to generate a video by learning the motion from the driving video while preserving the appearance from the source image. In this paper, we propose a novel structure-aware motion modeling approach, the deformable anchor model (DAM), which can automatically discover the motion structure of arbitrary objects without leveraging their prior structure information. Specifically, inspired by the known deformable part model (DPM), our DAM introduces two types of anchors or key-points: i) a number of motion anchors that capture both appearance and motion information from the source image and driving video; ii) a latent root anchor, which is linked to the motion anchors to facilitate better learning of the representations of the object structure information. More-over, DAM can be further extended to a hierarchical version through the introduction of additional latent anchors to model more complicated structures. By regularizing motion anchors with latent anchor(s), DAM enforces the corre-spondences between them to ensure the structural information is well captured and preserved. Moreover, DAM can be learned effectively in an unsupervised manner. We validate our proposed DAM for motion transfer on different bench-mark datasets. Extensive experiments clearly demonstrate that DAM achieves superior performance relative to existing state-of-the-art methods.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper15
- Text2Human: text-driven controllable human image generationYuming Jiang, Shuai Yang, Haonan Qiu, Wayne Wu 等SIGGRAPH 2022 · 被引用 140 次
- Implicit Identity Representation Conditioned Memory Compensation Network for Talking Head Video GenerationFa-Ting Hong, Dan XuICCV 2023 · 被引用 75 次
- Space-Time Diffusion Features for Zero-Shot Text-Driven Motion TransferDanah Yatim, Rafail Fridman, Omer Bar-Tal, Yoni Kasten 等CVPR 2024 · 被引用 29 次
- Wakey-Wakey: Animate Text by Mimicking Characters in a GIFLiwenhan Xie, Zhaoyu Zhou, Kerun Yu, Yun Wang 等UIST 2023 · 被引用 17 次
- VOODOO 3D: Volumetric Portrait Disentanglement for One-Shot 3D Head ReenactmentPhong Tran, Egor Zakharov, Long-Nhat Ho, Anh Tuan Tran 等CVPR 2024 · 被引用 15 次
它引用的顶会 Paper15
- Everybody Dance NowCaroline Chan, Shiry Ginosar, Tinghui Zhou, Alexei A. EfrosICCV 2019 · 被引用 840 次
- Liquid Warping GAN: A Unified Framework for Human Motion Imitation, Appearance Transfer and Novel View SynthesisWen Liu, Zhixin Piao, Jie Min, Wenhan Luo 等ICCV 2019 · 被引用 285 次
- FW-GAN: Flow-Navigated Warping GAN for Video Virtual Try-OnHaoye Dong, Xiaodan Liang, Xiaohui Shen, Bowen Wu 等ICCV 2019 · 被引用 130 次
- FLNet: Landmark Driven Fetching and Learning Network for Faithful Talking Facial Animation SynthesisKuangxiao Gu, Yuqian Zhou, Thomas S. HuangAAAI 2020 · 被引用 63 次
- One-shot Face Reenactment Using Appearance Adaptive NormalizationGuangming Yao, Yi Yuan, Tianjia Shao, Shuang Li 等AAAI 2021 · 被引用 30 次
相关 Paper
- Motion Representations for Articulated AnimationAliaksandr Siarohin, Oliver J. Woodford, Jian Ren, Menglei Chai 等CVPR 2021
- Learning Motion Refinement for Unsupervised Face AnimationJiale Tao, Shuhang Gu, Wen Li, Lixin DuanNeurIPS 2023 · 被引用 10 次
- Bidirectionally Deformable Motion Modulation For Video-based Human Pose TransferWing Yin Yu, Lai-Man Po, Ray C. C. Cheung, Yuzhi Zhao 等ICCV 2023 · 被引用 30 次
- Large Displacement Motion Transfer with Unsupervised Anytime InterpolationGuixiang Wang, Jianjun LiICML 2025
- Unpaired motion style transfer from video to animationKfir Aberman, Yijia Weng, Dani Lischinski, Daniel Cohen-Or 等SIGGRAPH 2020 · 被引用 178 次
