Learning Physics-Based Full-Body Human Reaching and Grasping from Brief Walking References
Yitang Li, Mingxian Lin, Zhuo Lin, Yipeng Deng, Yue Cao, Li Yi
Abstract
Existing motion generation methods based on MoCap data are often limited by data quality and coverage. In this work, we propose a framework that generates diverse, physically feasible full-body human reaching and grasping motions using only brief walking MoCap data. Based on the observation that walking data captures valuable movement patterns transferable across tasks and, on the other hand, the advanced kinematic methods can generate diverse grasping poses, which can then be interpolated into motions to serve as task-specific guidance. Our approach incorporates an active data generation strategy to maximize the utility of the generated motions, along with a local feature alignment mechanism that transfers natural movement patterns from walking data to enhance both the success rate and naturalness of the synthesized motions. By combining the fidelity and stability of natural walking with the flexibility and generalizability of task-specific generated data, our method demonstrates strong performance and robust adaptability in diverse scenes and with unseen objects.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers5
- CoDA: Coordinated Diffusion Noise Optimization for Whole-Body Manipulation of Articulated ObjectsHuaijin Pi, Zhi Cen, Zhiyang Dou, Taku KomuraNeurIPS 2025 · 14 citations
- InterPrior: Scaling Generative Control for Physics-Based Human-Object InteractionsSirui Xu, Samuel Schulter, Morteza Ziyadi, Xialin He et al.CVPR 2026 · 14 citations
- TeamHOI: Learning a Unified Policy for Cooperative Human-Object Interactions with Any Team SizeStefan Lionar, Gim Hee LeeCVPR 2026 · 3 citations
- Push-and-Step: From RL-Based Balance Recovery to Physical Simulation of Dense CrowdsAlexis Jensen, Pei Xu, Ioannis Karamouzas, Charles Pontonnier et al.CVPR 2026
- TokenHSI: Unified Synthesis of Physical Human-Scene Interactions through Task TokenizationLiang Pan, Zeshi Yang, Zhiyang Dou, Wenjia Wang et al.CVPR 2025
Builds on19
- AMASS: Archive of Motion Capture As Surface ShapesNaureen Mahmood, Nima Ghorbani, Nikolaus F. Troje, Gerard Pons-Moll et al.ICCV 2019 · 1,784 citations
- AMP: adversarial motion priors for stylized physics-based character controlXue Bin Peng, Ze Ma, Pieter Abbeel, Sergey Levine et al.SIGGRAPH 2021 · 392 citations
- Perpetual Humanoid Control for Real-time Simulated AvatarsZhengyi Luo, Jinkun Cao, Alexander Winkler, Kris Kitani et al.ICCV 2023 · 256 citations
- ASE: large-scale reusable adversarial skill embeddings for physically simulated charactersXue Bin Peng, Yunrong Guo, Lina Halper, Sergey Levine et al.SIGGRAPH 2022 · 217 citations
- AvatarCLIP: zero-shot text-driven generation and animation of 3D avatarsFangzhou Hong, Mingyuan Zhang, Liang Pan, Zhongang Cai et al.SIGGRAPH 2022 · 213 citations
Related papers
- COOP: Decoupling and Coupling of Whole-Body Grasping Pose GenerationYanzhao Zheng, Yunzhou Shi, Yuhao Cui, Zhongzhou Zhao et al.ICCV 2023 · 8 citations
- DiffGrasp: Whole-Body Grasping Synthesis Guided by Object Motion Using a Diffusion ModelYonghao Zhang, Qiang He, Yanguang Wan, Yinda Zhang et al.AAAI 2025 · 10 citations
- Hierarchical Generation of Human-Object Interactions with Diffusion Probabilistic ModelsHuaijin Pi, Sida Peng, Minghui Yang, Xiaowei Zhou et al.ICCV 2023 · 48 citations
- PMP: Learning to Physically Interact with Environments using Part-wise Motion PriorsJinseok Bae, Jungdam Won, Donggeun Lim, Cheol-Hui Min et al.SIGGRAPH 2023 · 27 citations
- D-Grasp: Physically Plausible Dynamic Grasp Synthesis for Hand-Object InteractionsSammy Joe Christen, Muhammed Kocabas, Emre Aksan, Jemin Hwangbo et al.CVPR 2022 · 69 citations
