Action-Conditioned Generation of Bimanual Object Manipulation Sequences
Haziq Razali, Yiannis Demiris
摘要
The generation of bimanual object manipulation sequences given a semantic action label has broad applications in collaborative robots or augmented reality. This relatively new problem differs from existing works that generate whole-body motions without any object interaction as it now requires the model to additionally learn the spatio-temporal relationship that exists between the human joints and object motion given said label. To tackle this task, we leverage the varying degree each muscle or joint is involved during object manipulation. For instance, the wrists act as the prime movers for the objects while the finger joints are angled to provide a firm grip. The remaining body joints are the least involved in that they are positioned as naturally and comfortably as possible. We thus design an architecture that comprises 3 main components: (i) a graph recurrent network that generates the wrist and object motion, (ii) an attention-based recurrent network that estimates the required finger joint angles given the graph configuration, and (iii) a recurrent network that reconstructs the body pose given the locations of the wrist. We evaluate our approach on the KIT Motion Capture and KIT RGBD Bimanual Manipulation datasets and show improvements over a simplified approach that treats the entire body as a single entity, and existing whole-body-only methods.
- Yiannis Demiris is supported by a Royal Academy of Engineering Chair in Emerging Technologies.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper7
- InterDiff: Generating 3D Human-Object Interactions with Physics-Informed DiffusionSirui Xu, Zhengyuan Li, Yu-Xiong Wang, Liang-Yan GuiICCV 2023 · 被引用 201 次
- InterDreamer: Zero-Shot Text to 3D Dynamic Human-Object InteractionSirui Xu, Ziyin Wang, Yu-Xiong Wang, Liangyan GuiNeurIPS 2024 · 被引用 78 次
- Forecasting Bimanual Object Manipulation Sequences from Unimanual ObservationsHaziq Razali, Yiannis DemirisAAAI 2024 · 被引用 1 次
- Learning Diverse Bimanual Dexterous Manipulation Skills from Human DemonstrationsBohan Zhou, Haoqi Yuan, Yuhui Fu, Zongqing LuAAAI 2026
- InterMimic: Towards Universal Whole-Body Control for Physics-Based Human-Object InteractionsSirui Xu, Hung Yu Ling, Yu-Xiong Wang, Liang-Yan GuiCVPR 2025
它引用的顶会 Paper7
- AI Choreographer: Music Conditioned 3D Dance Generation with AIST++Ruilong Li, Shan Yang, David A. Ross, Angjoo KanazawaICCV 2021 · 被引用 701 次
- Action-Conditioned 3D Human Motion Synthesis with Transformer VAEMathis Petrovich, Michael J. Black, Gül VarolICCV 2021 · 被引用 672 次
- Action2Motion: Conditioned Generation of 3D Human MotionsChuan Guo, Xinxin Zuo, Sen Wang, Shihao Zou 等ACM MM 2020 · 被引用 394 次
- GOAL: Generating 4D Whole-Body Motion for Hand-Object GraspingOmid Taheri, Vasileios Choutas, Michael J. Black, Dimitrios TzionasCVPR 2022 · 被引用 103 次
- Aggregated Multi-GANs for Controlled 3D Human Motion PredictionZhenguang Liu, Kedi Lyu, Shuang Wu, Haipeng Chen 等AAAI 2021 · 被引用 64 次
相关 Paper
- DiffGrasp: Whole-Body Grasping Synthesis Guided by Object Motion Using a Diffusion ModelYonghao Zhang, Qiang He, Yanguang Wan, Yinda Zhang 等AAAI 2025 · 被引用 10 次
- ManipNet: neural manipulation synthesis with a hand-object spatial representationHe Zhang, Yuting Ye, Takaaki Shiratori, Taku KomuraSIGGRAPH 2021 · 被引用 70 次
- BimArt: A Unified Approach for the Synthesis of 3D Bimanual Interaction with Articulated ObjectsWanyue Zhang, Rishabh Dabral, Vladislav Golyanik, Vasileios Choutas 等CVPR 2025
- Context-Aware Human Motion PredictionEnric Corona, Albert Pumarola, Guillem Alenyà, Francesc Moreno-NoguerCVPR 2020
- Deep Imitation Learning for Bimanual Robotic ManipulationFan Xie, Alexander Chowdhury, M. Clara De Paolis Kaluza, Linfeng Zhao 等NeurIPS 2020 · 被引用 108 次
