A Unified Framework for Real Time Motion Completion
Yinglin Duan, Yue Lin, Zhengxia Zou, Yi Yuan, Zhehui Qian, Bohan Zhang
Abstract
Motion completion, as a challenging and fundamental problem, is of great significance in film and game applications. For different motion completion application scenarios (in-betweening, in-filling, and blending), most previous methods deal with the completion problems with case-by-case methodology designs. In this work, we propose a simple but effective method to solve multiple motion completion problems under a unified framework and achieves a new state-of-the-art accuracy on LaFAN1 (+17% better than previous sota) under multiple evaluation settings. Inspired by the recent great success of self-attention-based transformer models, we consider the completion as a sequence-to-sequence prediction problem. Our method consists of three modules - a standard transformer encoder with self-attention that learns long-range dependencies of input motions, a trainable mixture embedding module that models temporal information and encodes different key-frame combinations in a unified form, and a new motion perceptual loss for better capturing high-frequency movements. Our method can predict multiple missing frames within a single forward propagation in real-time and get rid of the post-processing requirement. We also introduce a novel large-scale dance movement dataset for exploring the scaling capability of our method and its effectiveness in complex motion applications.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 10dcb2b6-64a7-4cac-b7de-11e766969b44Cited by top-tier papers10
- FLAME: Free-Form Language-Based Motion Synthesis & EditingJihoon Kim, Jiseob Kim, Sungjoon ChoiAAAI 2023 · 276 citations
- Example-based Motion Synthesis via Generative Motion MatchingWeiyu Li, Xuelin Chen, Peizhuo Li, Olga Sorkine-Hornung et al.SIGGRAPH 2023 · 24 citations
- Skeleton2Humanoid: Animating Simulated Characters for Physically-plausible Motion In-betweeningYunhao Li, Zhenbo Yu, Yucheng Zhu, Bingbing Ni et al.ACM MM 2022 · 8 citations
- COOP: Decoupling and Coupling of Whole-Body Grasping Pose GenerationYanzhao Zheng, Yunzhou Shi, Yuhao Cui, Zhongzhou Zhao et al.ICCV 2023 · 8 citations
- MixSynthFormer: A Transformer Encoder-like Structure with Mixed Synthetic Self-attention for Efficient Human Pose EstimationYuran Sun, Alan William Dougherty, Zhuoying Zhang, Yi-King Choi et al.ICCV 2023 · 6 citations
Builds on7
- Learning to Reconstruct 3D Human Pose and Shape via Model-Fitting in the LoopNikos Kolotouros, Georgios Pavlakos, Michael J. Black, Kostas DaniilidisICCV 2019 · 1,139 citations
- AI Choreographer: Music Conditioned 3D Dance Generation with AIST++Ruilong Li, Shan Yang, David A. Ross, Angjoo KanazawaICCV 2021 · 701 citations
- Stabilizing Transformers for Reinforcement LearningEmilio Parisotto, H. Francis Song, Jack W. Rae, Razvan Pascanu et al.ICML 2020 · 464 citations
- Robust motion in-betweeningFélix G. Harvey, Mike Yurick, Derek Nowrouzezahrai, Christopher J. PalSIGGRAPH 2020 · 269 citations
- Human Motion Prediction via Spatio-Temporal InpaintingAlejandro Hernandez Ruiz, Jürgen Gall, Francesc MorenoICCV 2019 · 233 citations
Related papers
- A Dual-Masked Auto-Encoder for Robust Motion Capture with Spatial-Temporal Skeletal Token CompletionJunkun Jiang, Jie Chen, Yike GuoACM MM 2022 · 8 citations
- A Unified Masked Autoencoder with Patchified Skeletons for Motion SynthesisEsteve Valls Mascaro, Hyemin Ahn, Dongheui LeeAAAI 2024 · 11 citations
- A Unified 3D Human Motion Synthesis Model via Conditional Variational Auto-Encoder∗Yujun Cai, Yiwei Wang, Yiheng Zhu, Tat-Jen Cham et al.ICCV 2021 · 83 citations
- Continuous Intermediate Token Learning with Implicit Motion Manifold for Keyframe Based Motion InterpolationClinton Ansun Mo, Kun Hu, Chengjiang Long, Zhiyong WangCVPR 2023
- Frequency-Aware Spatiotemporal Transformers for Video Inpainting DetectionBingyao Yu, Wanhua Li, Xiu Li, Jiwen Lu et al.ICCV 2021 · 38 citations
