MetaPix: Few-Shot Video Retargeting
Jessica Lee, Deva Ramanan, Rohit Girdhar
Abstract
We address the task of unsupervised retargeting of human actions from one video to another. We consider the challenging setting where only a few frames of the target is available. The core of our approach is a conditional generative model that can transcode input skeletal poses (automatically extracted with an off-the-shelf pose estimator) to output target frames. However, it is challenging to build a universal transcoder because humans can appear wildly different due to clothing and background scene geometry. Instead, we learn to adapt - or personalize - a universal generator to the particular human and background in the target. To do so, we make use of meta-learning to discover effective strategies for on-the-fly personalization. One significant benefit of meta-learning is that the personalized transcoder naturally enforces temporal coherence across its generated frames; all frames contain consistent clothing and background geometry of the target. We experiment on in-the-wild internet videos and images and show our approach improves over widely-used baselines for the task.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers4
- Disco: Disentangled Control for Realistic Human Dance GenerationTan Wang, Linjie Li, Kevin Lin, Yuanhao Zhai et al.CVPR 2024 · 62 citations
- Harnessing Meta-Learning for Improving Full-Frame Video StabilizationMuhammad Kashif Ali, Eun Woo Im, Dongjin Kim, Tae Hyun KimCVPR 2024
- Flow Guided Transformable Bottleneck Networks for Motion RetargetingJian Ren, Menglei Chai, Oliver J. Woodford, Kyle Olszewski et al.CVPR 2021
- Few-Shot Human Motion Transfer by Personalized Geometry and Texture ModelingZhichao Huang, Xintong Han, Jia Xu, Tong ZhangCVPR 2021
Builds on2
Related papers
- Scene-Adaptive Video Frame Interpolation via Meta-LearningMyungsub Choi, Janghoon Choi, Sungyong Baik, Tae Hyun Kim et al.CVPR 2020
- MoCaNet: Motion Retargeting In-the-Wild via Canonicalization NetworksWentao Zhu, Zhuoqian Yang, Ziang Di, Wayne Wu et al.AAAI 2022 · 24 citations
- Point-Based Modeling of Human ClothingIlya Zakharkin, Kirill Mazur, Artur Grigorev, Victor LempitskyICCV 2021 · 53 citations
- Learning Motion-Dependent Appearance for High-Fidelity Rendering of Dynamic Humans from a Single CameraJae Shin Yoon, Duygu Ceylan, Tuanfeng Y. Wang, Jingwan Lu et al.CVPR 2022 · 10 citations
- Neural Marionette: Unsupervised Learning of Motion Skeleton and Latent Dynamics from Volumetric VideoJinseok Bae, Hojun Jang, Cheol-Hui Min, Hyungun Choi et al.AAAI 2022 · 6 citations
