Bidirectionally Deformable Motion Modulation For Video-based Human Pose Transfer
Wing Yin Yu, Lai-Man Po, Ray C. C. Cheung, Yuzhi Zhao, Yu Xue, Kun Li
Abstract
Video-based human pose transfer is a video-to-video generation task that animates a plain source human image based on a series of target human poses. Considering the difficulties in transferring highly structural patterns on the garments and discontinuous poses, existing methods often generate unsatisfactory results such as distorted textures and flickering artifacts. To address these issues, we propose a novel Deformable Motion Modulation (DMM) that utilizes geometric kernel offset with adaptive weight modulation to simultaneously perform feature alignment and style transfer. Different from normal style modulation used in style transfer, the proposed modulation mechanism adaptively reconstructs smoothed frames from style codes according to the object shape through an irregular receptive field of view. To enhance the spatio-temporal consistency, we leverage bidirectional propagation to extract the hidden motion information from a warped image sequence generated by noisy poses. The proposed feature propagation significantly enhances the motion prediction ability by forward and backward propagation. Both quantitative and qualitative experimental results demonstrate superiority over the state-of-the-arts in terms of image fidelity and visual continuity. The source code is publicly available at github.com/rocketappslab/bdmm.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 0fb96416-429d-45e9-bb86-dc2591f5da35Cited by top-tier papers6
- MagicFight: Personalized Martial Arts Combat Video GenerationJiancheng Huang, Mingfu Yan, Songyan Chen, Yi Huang et al.ACM MM 2024 · 16 citations
- Animate Anyone 2: High-Fidelity Character Image Animation with Environment AffordanceLi Hu, Guangyuan Wang, Zhen Shen, Xin Gao et al.ICCV 2025 · 6 citations
- DreamDance: Animating Human Images by Enriching 3D Geometry Cues from 2D PosesYatian Pang, Bin Zhu, Bin Lin, Mingzhe Zheng et al.ICCV 2025 · 2 citations
- Multi-Identity Human Image Animation with Structural Video DiffusionZhenzhi Wang, Yixuan Li, Yanhong Zeng, Yuwei Guo et al.ICCV 2025 · 1 citation
- Animate-X: Universal Character Image Animation with Enhanced Motion RepresentationShuai Tan, Biao Gong, Xiang Wang, Shiwei Zhang et al.ICLR 2025
Builds on11
- Learning to Reconstruct 3D Human Pose and Shape via Model-Fitting in the LoopNikos Kolotouros, Georgios Pavlakos, Michael J. Black, Kostas DaniilidisICCV 2019 · 1,139 citations
- BasicVSR++: Improving Video Super-Resolution with Enhanced Propagation and AlignmentKelvin C. K. Chan, Shangchen Zhou, Xiangyu Xu, Chen Change LoyCVPR 2022 · 522 citations
- Liquid Warping GAN: A Unified Framework for Human Motion Imitation, Appearance Transfer and Novel View SynthesisWen Liu, Zhixin Piao, Jie Min, Wenhan Luo et al.ICCV 2019 · 285 citations
- Towards Multi-Pose Guided Virtual Try-On NetworkHaoye Dong, Xiaodan Liang, Xiaohui Shen, Bochao Wang et al.ICCV 2019 · 226 citations
- Towards An End-to-End Framework for Flow-Guided Video InpaintingZhen Li, Chengze Lu, Jianhua Qin, Chun-Le Guo et al.CVPR 2022 · 136 citations
Related papers
- High-Fidelity Neural Human Motion Transfer From Monocular VideoMoritz Kappel, Vladislav Golyanik, Mohamed A. Elgharib, Jann-Ole Henningson et al.CVPR 2021
- MoST: Motion Style Transformer Between Diverse Action ContentsBoeun Kim, Jungho Kim, Hyung Jin Chang, Jin Young ChoiCVPR 2024
- Structure-Aware Motion Transfer with Deformable Anchor ModelJiale Tao, Biao Wang, Borun Xu, Tiezheng Ge et al.CVPR 2022 · 33 citations
- Dual Conditioned Motion Diffusion for Pose-Based Video Anomaly DetectionHongsong Wang, Andi Xu, Pinle Ding, Jie GuiAAAI 2025 · 8 citations
- Large Displacement Motion Transfer with Unsupervised Anytime InterpolationGuixiang Wang, Jianjun LiICML 2025
