WalkTheDog: Cross-Morphology Motion Alignment via Phase Manifolds
Peizhuo Li, Sebastian Starke, Yuting Ye, Olga Sorkine-Hornung
Abstract
We present a new approach for understanding the periodicity structure and semantics of motion datasets, independently of the morphology and skeletal structure of characters. Unlike existing methods using an overly sparse high-dimensional latent, we propose a phase manifold consisting of multiple closed curves, each corresponding to a latent amplitude. With our proposed vector quantized periodic autoencoder, we learn a shared phase manifold for multiple characters, such as a human and a dog, without any supervision. This is achieved by exploiting the discrete structure and a shallow network as bottlenecks, such that semantically similar motions are clustered into the same curve of the manifold, and the motions within the same component are aligned temporally by the phase variable. In combination with an improved motion matching framework, we demonstrate the manifold’s capability of timing and semantics alignment in several applications, including motion retrieval, transfer and stylization. Code and pre-trained models for this paper are available at peizhuoli.github.io/walkthedog.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 8b5f442f-9ab6-4ce5-8445-03bd01faf57cCited by top-tier papers6
- AnyTop: Character Animation Diffusion with Any TopologyInbar Gat, Sigal Raab, Guy Tevet, Yuval Reshef et al.SIGGRAPH 2025 · 9 citations
- Semantic-Aware Motion Encoding for Topology-Agnostic Character AnimationZongye Zhang, Yuzhuo Cui, Qingjie Liu, Yunhong WangICML 2026 · 1 citation
- FunPhase: A Periodic Functional Autoencoder for Motion Generation via Phase ManifoldsMarco Pegoraro, Evan Atherton, Bruno Roy, Aliasghar Khani et al.ICML 2026 · 1 citation
- STyMo: Fast and Controllable Few-Shot Motion Style TransferJose Luis Ponton, Alexander W. Winkler, Ladislav Kavan, Yuting Ye et al.SIGGRAPH 2026
- POMP: Physics-constrainable Motion Generative Model through Phase ManifoldsBin Ji, Ye Pan, Zhimeng Liu, Shuai Tan et al.CVPR 2025
Builds on10
- High-Resolution Image Synthesis with Latent Diffusion ModelsRobin Rombach, Andreas Blattmann, Dominik Lorenz, Patrick Esser et al.CVPR 2022 · 13,123 citations
- Skeleton-aware networks for deep motion retargetingKfir Aberman, Peizhuo Li, Dani Lischinski, Olga Sorkine-Hornung et al.SIGGRAPH 2020 · 210 citations
- Local motion phases for learning multi-contact character movementsSebastian Starke, Yiwei Zhao, Taku Komura, Kazi A. ZamanSIGGRAPH 2020 · 186 citations
- Unpaired motion style transfer from video to animationKfir Aberman, Yijia Weng, Dani Lischinski, Daniel Cohen-Or et al.SIGGRAPH 2020 · 178 citations
- Bailando: 3D Dance Generation by Actor-Critic GPT with Choreographic MemoryLi Siyao, Weijiang Yu, Tianpei Gu, Chunze Lin et al.CVPR 2022 · 170 citations
Related papers
- DeepPhase: periodic autoencoders for learning motion phase manifoldsSebastian Starke, Ian Mason, Taku KomuraSIGGRAPH 2022 · 142 citations
- PUMPS: Skeleton-Agnostic Point-Based Universal Motion Pre-Training for Synthesis in Human Motion TasksClinton Ansun Mo, Kun Hu, Chengjiang Long, Dong Yuan et al.ICCV 2025 · 2 citations
- QPGesture: Quantization-Based and Phase-Guided Motion Matching for Natural Speech-Driven Gesture GenerationSicheng Yang, Zhiyong Wu, Minglei Li, Zhensong Zhang et al.CVPR 2023
- PhaseMP: Robust 3D Pose Estimation via Phase-conditioned Human Motion PriorMingyi Shi, Sebastian Starke, Yuting Ye, Taku Komura et al.ICCV 2023 · 27 citations
- Deep Compositional Phase Diffusion for Long Motion Sequence GenerationHo Yin Au, Jie Chen, Junkun Jiang, Jingyu XiangNeurIPS 2025 · 7 citations
