Neural Marionette: Unsupervised Learning of Motion Skeleton and Latent Dynamics from Volumetric Video
Jinseok Bae, Hojun Jang, Cheol-Hui Min, Hyungun Choi, Young Min Kim
摘要
We present Neural Marionette, an unsupervised approach that discovers the skeletal structure from a dynamic sequence and learns to generate diverse motions that are consistent with the observed motion dynamics. Given a video stream of point cloud observation of an articulated body under arbitrary motion, our approach discovers the unknown low-dimensional skeletal relationship that can effectively represent the movement. Then the discovered structure is utilized to encode the motion priors of dynamic sequences in a latent structure, which can be decoded to the relative joint rotations to represent the full skeletal motion. Our approach works without any prior knowledge of the underlying motion or skeletal structure, and we demonstrate that the discovered structure is even comparable to the hand-labeled ground truth skeleton in representing a 4D sequence of motion. The skeletal structure embeds the general semantics of possible motion space that can generate motions for diverse scenarios. We verify that the learned motion prior is generalizable to the multi-modal sequence generation, interpolation of two poses, and motion retargeting to a different skeletal structure.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper3
- Dynamic Mesh Recovery from Partial Point Cloud SequenceHojun Jang, Minkwan Kim, Jinseok Bae, Young Min KimICCV 2023 · 被引用 5 次
- Transfer4D: A Framework for Frugal Motion Capture and Deformation TransferShubh Maheshwari, Rahul Narain, Ramya HebbalaguppeCVPR 2023
- RigGS: Rigging of 3D Gaussians for Modeling Articulated Objects in VideosYuxin Yao, Zhi Deng, Junhui HouCVPR 2025
它引用的顶会 Paper14
- Learning Trajectory Dependencies for Human Motion PredictionWei Mao, Miaomiao Liu, Mathieu Salzmann, Hongdong LiICCV 2019 · 被引用 534 次
- FreiHAND: A Dataset for Markerless Capture of Hand Pose and Shape From Single RGB ImagesChristian Zimmermann, Duygu Ceylan, Jimei Yang, Bryan C. Russell 等ICCV 2019 · 被引用 493 次
- GENESIS: Generative Scene Inference and Sampling with Object-Centric Latent RepresentationsMartin Engelcke, Adam R. Kosiorek, Oiwi Parker Jones, Ingmar PosnerICLR 2020 · 被引用 334 次
- Causal Discovery in Physical Systems from VideosYunzhu Li, Antonio Torralba, Anima Anandkumar, Dieter Fox 等NeurIPS 2020 · 被引用 133 次
- NPMs: Neural Parametric Models for 3D Deformable ShapesPablo R. Palafox, Aljaz Bozic, Justus Thies, Matthias Nießner 等ICCV 2021 · 被引用 129 次
相关 Paper
- DeepPhase: periodic autoencoders for learning motion phase manifoldsSebastian Starke, Ian Mason, Taku KomuraSIGGRAPH 2022 · 被引用 142 次
- RigMo: Unifying Rig and Motion Learning for Generative AnimationHao Zhang, Jiahao Luo, Bohui Wan, Yizhou Zhao 等CVPR 2026 · 被引用 6 次
- Nonparametric Object and Parts Modeling With Lie Group DynamicsDavid S. Hayden, Jason Pacheco, John W. Fisher IIICVPR 2020
- Learning Diverse Stochastic Human-Action Generators by Learning Smooth Latent TransitionsZhenyi Wang, Ping Yu, Yang Zhao, Ruiyi Zhang 等AAAI 2020 · 被引用 73 次
- DIMO: Diverse 3D Motion Generation for Arbitrary ObjectsLinzhan Mou, Jiahui Lei, Chen Wang, Lingjie Liu 等ICCV 2025 · 被引用 2 次
