Motion Synthesis with Sparse and Flexible Keyjoint Control
Inwoo Hwang, Jinseok Bae, Donggeun Lim, Young Min Kim
Abstract
Creating expressive character animations is laborintensive, requiring intricate manual adjustment of animators across space and time. Previous works on controllable motion generation often rely on a predefined set of dense spatio-temporal specifications (e.g., dense pelvis trajectories with exact per-frame timing), limiting practicality for animators. To process high-level intent and intuitive control in diverse scenarios, we propose a practical controllable motions synthesis framework that respects sparse and flexible keyjoint signals. Our approach employs a decomposed diffusion-based motion synthesis framework that first synthesizes keyjoint movements from sparse input control signals † Corresponding author and then synthesizes full-body motion based on the completed keyjoint trajectories. The low-dimensional keyjoint movements can easily adapt to various control signal types, such as end-effector position for diverse goal-driven motion synthesis, or incorporate functional constraints on a subset of keyjoints. Additionally, we introduce a time-agnostic control formulation, eliminating the need for frame-specific timing annotations and enhancing control flexibility. Then, the shared second stage can synthesize a natural whole-body motion that precisely satisfies the task requirement from dense keyjoint movements. We demonstrate the effectiveness of sparse and flexible keyjoint control through comprehensive experiments on diverse datasets and scenarios. Project
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext c6e66ffe-ff7d-4151-8987-60b382c27570Cited by top-tier papers3
- Pressure2Motion: Hierarchical Human Motion Reconstruction from Ground Pressure with Text GuidanceZhengxuan Li, Qinhui Yang, Yiyu Zhuang, Chuan Guo et al.CVPR 2026 · 1 citation
- Unifying Precise Keyframes and Semantic Control via Multi-level DiffusionLinjun Wu, Jiejia Yu, Leyang Jin, He Wang et al.CVPR 2026 · 1 citation
- Scenemi: Motion In-Betweening for Modeling Human-Scene InteractionsInwoo Hwang, Bing Zhou, Young Min Kim, Jian Wang et al.ICCV 2025
Builds on42
- Denoising Diffusion Probabilistic ModelsJonathan Ho, Ajay Jain, Pieter AbbeelNeurIPS 2020 · 35,902 citations
- Directly Denoising Diffusion ModelsDan Zhang, Jingjing Wang, Feng LuoICML 2024 · 11,724 citations
- Adding Conditional Control to Text-to-Image Diffusion ModelsLvmin Zhang, Anyi Rao, Maneesh AgrawalaICCV 2023 · 6,759 citations
- Improved Denoising Diffusion Probabilistic ModelsAlexander Quinn Nichol, Prafulla DhariwalICML 2021 · 5,234 citations
- AI Choreographer: Music Conditioned 3D Dance Generation with AIST++Ruilong Li, Shan Yang, David A. Ross, Angjoo KanazawaICCV 2021 · 701 citations
Related papers
- AutoKeyframe: Autoregressive Keyframe Generation for Human Motion Synthesis and EditingBowen Zheng, Ke Chen, Yuxin Yao, Zijiao Zeng et al.SIGGRAPH 2025 · 3 citations
- DisPose: Disentangling Pose Guidance for Controllable Human Image AnimationHongxiang Li, Yaowei Li, Yuhang Yang, Junjie Cao et al.ICLR 2025
- PMG: Progressive Motion Generation via Sparse Anchor Postures Curriculum LearningYingjie Xi, Jian Jun Zhang, Xiaosong YangACM MM 2025 · 1 citation
- Enhanced Fine-Grained Motion Diffusion for Text-Driven Human Motion SynthesisDong Wei, Xiaoning Sun, Huaijiang Sun, Shengxiang Hu et al.AAAI 2024 · 13 citations
- Less is More: Improving Motion Diffusion Models with Sparse KeyframesJinseok Bae, Inwoo Hwang, Young Yoon Lee, Ziyu Guo et al.ICCV 2025 · 4 citations
