Towards Efficient and Diverse Generative Model for Unconditional Human Motion Synthesis
Hua Yu, Weiming Liu, Jiapeng Bai, Xu Gui, Yaqing Hou, Yew-Soon Ong, Qiang Zhang
摘要
Recent generative methods have revolutionized the way of human motion synthesis, such as Variational Autoencoders (VAEs), Generative Adversarial Networks (GANs), and Denoising Diffusion Probabilistic Models (DMs). These methods have gained significant attention in human motion fields. However, there are still challenges in unconditionally generating highly diverse human motions from a given distribution. To enhance the diversity of synthesized human motions, previous methods usually employ deep neural networks (DNNs) to train a transport map that transforms Gaussian noise distribution into real human motion distribution. According to Figalli's regularity theory, the optimal transport map computed by DNNs frequently exhibits discontinuities. This is due to the inherent limitation of DNNs in representing only continuous maps. Consequently, the generated human motions tend to heavily concentrate on densely populated regions of the data distribution, resulting in mode collapse or mode mixture. To address the issues, we propose an efficient method called MOOT for unconditional human motion synthesis. First, we utilize a reconstruction network based on GRU and transformer to map human motions to latent space. Next, we employ convex optimization to match the noise distribution with the latent space distribution of human motions through the Optimal Transport (OT) map. Then, we combine the extended OT map with the generator of reconstruction network to generate new human motions. Thereby overcoming the issues of mode collapse and mode mixture. MOOT generates a latent code distribution that is well-behaved and highly structured, providing a strong motion prior for various applications in the field of human motion. Through qualitative and quantitative experiments, MOOT achieves state-of-the-art results surpassing the latest methods, validating its superiority in unconditional human motion generation.
问问这篇 Paper
问问你的智能体。
Lune 读过与它相关的顶会 Paper,每个回答都会注明依据哪几篇。
引用它的顶会 Paper3
- Geometric Neural Distance Fields for Learning Human Motion PriorsZhengdi Yu, Simone Foti, Linguang Zhang, Amy Zhao 等CVPR 2026 · 被引用 8 次
- PoseD-Flow: Versatile and Guided Flow Matching Model of Human PoseJebastin Nadar, Simone Foti, Tolga BirdalCVPR 2026 · 被引用 3 次
- Deterministic-to-Stochastic Diverse Latent Feature Mapping for Human Motion SynthesisHua Yu, Weiming Liu, Gui Xu, Yaqing Hou 等CVPR 2025
相关 Paper
- Ae-OT: a New Generative Model based on Extended Semi-discrete Optimal transportDongsheng An, Yang Guo, Na Lei, Zhongxuan Luo 等ICLR 2020 · 被引用 68 次
- MoDi: Unconditional Motion Synthesis from Diverse DataSigal Raab, Inbal Leibovitch, Peizhuo Li, Kfir Aberman 等CVPR 2023
- Generating Smooth Pose Sequences for Diverse Human Motion PredictionWei Mao, Miaomiao Liu, Mathieu SalzmannICCV 2021 · 被引用 101 次
- Executing your Commands via Motion Diffusion in Latent SpaceXin Chen, Biao Jiang, Wen Liu, Zilong Huang 等CVPR 2023
- Motion Diversification NetworksHee Jae Kim, Eshed Ohn-BarCVPR 2024
