BimArt: A Unified Approach for the Synthesis of 3D Bimanual Interaction with Articulated Objects
Wanyue Zhang, Rishabh Dabral, Vladislav Golyanik, Vasileios Choutas, Eduardo Alvarado, Thabo Beeler, Marc Habermann, Christian Theobalt
摘要
We present BimArt, a novel generative approach for synthesizing 3D bimanual hand interactions with articulated objects. Unlike prior works, we do not rely on a reference grasp, a coarse hand trajectory, or separate modes for grasping and articulating. To achieve this, we first generate distance-based contact maps conditioned on the object trajectory with an articulation-aware feature representation, revealing rich bimanual patterns for manipulation. The learned contact prior is then used to guide our hand motion generator, producing diverse and realistic bimanual motions for object movement and articulation. Our work offers key insights into feature representation and contact prior for articulated objects, demonstrating their effectiveness in taming the complex, high-dimensional space of bimanual hand-object interactions. Through comprehensive quantitative experiments, we demonstrate a clear step towards simplified and high-quality hand-object animations that surpass the state of the art in motion quality and diversity. Project page: https://vcai.mpi-inf.mpg.de/projects/bimart/.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper8
- OpenHOI: Open-World Hand-Object Interaction Synthesis with Multimodal Large Language ModelZhenhao Zhang, Ye Shi, Lingxiao Yang, Suting Ni 等NeurIPS 2025 · 被引用 25 次
- CoDA: Coordinated Diffusion Noise Optimization for Whole-Body Manipulation of Articulated ObjectsHuaijin Pi, Zhi Cen, Zhiyang Dou, Taku KomuraNeurIPS 2025 · 被引用 14 次
- Unleashing Guidance Without Classifiers for Human-Object Interaction AnimationZiyin Wang, Sirui Xu, Chuan Guo, Bing Zhou 等ICLR 2026 · 被引用 6 次
- CLUTCH: Contextualized Language model for Unlocking Text-Conditioned Hand motion modelling in the wildBalamurugan Thambiraja, Omid Taheri, Radek Danecek, Giorgio Becherini 等ICLR 2026 · 被引用 2 次
- RoMo: A Large-Scale, Richly Organized Dataset and Semantic Taxonomy for Human Motion GenerationJiahao Zhang, Joseph Liu, Young-Yoon Lee, Seonghyeon Moon 等CVPR 2026 · 被引用 2 次
它引用的顶会 Paper36
- Denoising Diffusion Probabilistic ModelsJonathan Ho, Ajay Jain, Pieter AbbeelNeurIPS 2020 · 被引用 35,902 次
- Denoising Diffusion Implicit ModelsJiaming Song, Chenlin Meng, Stefano ErmonICLR 2021 · 被引用 11,743 次
- AMASS: Archive of Motion Capture As Surface ShapesNaureen Mahmood, Nima Ghorbani, Nikolaus F. Troje, Gerard Pons-Moll 等ICCV 2019 · 被引用 1,784 次
- Action-Conditioned 3D Human Motion Synthesis with Transformer VAEMathis Petrovich, Michael J. Black, Gül VarolICCV 2021 · 被引用 672 次
- 3D-LLM: Injecting the 3D World into Large Language ModelsYining Hong, Haoyu Zhen, Peihao Chen, Shuhong Zheng 等NeurIPS 2023 · 被引用 662 次
相关 Paper
- ContactGen: Generative Contact Modeling for Grasp GenerationShaowei Liu, Yang Zhou, Jimei Yang, Saurabh Gupta 等ICCV 2023 · 被引用 60 次
- DiffGrasp: Whole-Body Grasping Synthesis Guided by Object Motion Using a Diffusion ModelYonghao Zhang, Qiang He, Yanguang Wan, Yinda Zhang 等AAAI 2025 · 被引用 10 次
- Decoupled Generative Modeling for Human-Object Interaction SynthesisHwanhee Jung, Seunggwan Lee, Jeongyoon Yoon, SeungHyeon Kim 等CVPR 2026 · 被引用 4 次
- HandX: Scaling Bimanual Motion and Interaction GenerationZimu Zhang, Yucheng Zhang, Xiyan Xu, Ziyin Wang 等CVPR 2026 · 被引用 2 次
- NAP: Neural 3D Articulated Object PriorJiahui Lei, Congyue Deng, William B. Shen, Leonidas J. Guibas 等NeurIPS 2023 · 被引用 53 次
