DualAfford: Learning Collaborative Visual Affordance for Dual-gripper Manipulation
Yan Zhao, Ruihai Wu, Zhehuan Chen, Yourong Zhang, Qingnan Fan, Kaichun Mo, Hao Dong
摘要
It is essential yet challenging for future home-assistant robots to understand and manipulate diverse 3D objects in daily human environments. Towards building scalable systems that can perform diverse manipulation tasks over various 3D shapes, recent works have advocated and demonstrated promising results learning visual actionable affordance, which labels every point over the input 3D geometry with an action likelihood of accomplishing the downstream task (e.g., pushing or picking-up). However, these works only studied single-gripper manipulation tasks, yet many real-world tasks require two hands to achieve collaboratively. In this work, we propose a novel learning framework, DualAfford, to learn collaborative affordance for dual-gripper manipulation tasks. The core design of the approach is to reduce the quadratic problem for two grippers into two disentangled yet interconnected subtasks for efficient learning. Using the large-scale PartNet-Mobility and ShapeNet datasets, we set up four benchmark tasks for dual-gripper manipulation. Experiments prove the effectiveness and superiority of our method over three baselines.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了最后一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper8
- LEMON: Learning 3D Human-Object Interaction Relation from 2D ImagesYuhang Yang, Wei Zhai, Hongchen Luo, Yang Cao 等CVPR 2024 · 被引用 12 次
- Seeing the Unseen: Visual Common Sense for Semantic PlacementRam Ramrakhya, Aniruddha Kembhavi, Dhruv Batra, Zsolt Kira 等CVPR 2024 · 被引用 3 次
- VLBiMan: Vision-Language Anchored One-Shot Demonstration Enables Generalizable Bimanual Robotic ManipulationHuayi Zhou, Kui JiaICLR 2026 · 被引用 3 次
- BiPreManip: Learning Affordance-Based Bimanual Preparatory Manipulation through Anticipatory CollaborationYan Shen, Feng Jiang, Zichen He, Xiaoqi Li 等CVPR 2026 · 被引用 3 次
- PA3FF: Learning Part-Aware Dense 3D Feature Field For Generalizable Articulated Object ManipulationYue Chen, Muqing Jiang, Kaifeng Zheng, Jiaqi Liang 等ICLR 2026 · 被引用 2 次
它引用的顶会 Paper9
- 6-DOF GraspNet: Variational Grasp Generation for Object ManipulationArsalan Mousavian, Clemens Eppner, Dieter FoxICCV 2019 · 被引用 673 次
- Hand-Object Contact Consistency Reasoning for Human Grasps GenerationHanwen Jiang, Shaowei Liu, Jiashun Wang, Xiaolong WangICCV 2021 · 被引用 242 次
- Where2Act: From Pixels to Actions for Articulated 3D ObjectsKaichun Mo, Leonidas J. Guibas, Mustafa Mukadam, Abhinav Gupta 等ICCV 2021 · 被引用 240 次
- VAT-Mart: Learning Visual Action Trajectory Proposals for Manipulating 3D ARTiculated ObjectsRuihai Wu, Yan Zhao, Kaichun Mo, Zizheng Guo 等ICLR 2022 · 被引用 119 次
- Deep Imitation Learning for Bimanual Robotic ManipulationFan Xie, Alexander Chowdhury, M. Clara De Paolis Kaluza, Linfeng Zhao 等NeurIPS 2020 · 被引用 108 次
相关 Paper
- 3D AffordanceNet: A Benchmark for Visual Object Affordance UnderstandingShengheng Deng, Xun Xu, Chaozheng Wu, Ke Chen 等CVPR 2021
- MAAL: Multimodality-Aware Autoencoder-based Affordance Learning for 3D Articulated ObjectsYuanzhi Liang, Xiaohan Wang, Linchao Zhu, Yi YangICCV 2023 · 被引用 3 次
- A3D: Adaptive Affordance Assembly with Dual-Arm ManipulationJiaqi Liang, Yue Chen, Qize Yu, Yan Shen 等AAAI 2026 · 被引用 3 次
- Learning Environment-Aware Affordance for 3D Articulated Object Manipulation under OcclusionsRuihai Wu, Kai Cheng, Yan Zhao, Chuanruo Ning 等NeurIPS 2023 · 被引用 43 次
- InteractMove: Text-Controlled Human-Object Interaction Generation in 3D Scenes with Movable ObjectsXinhao Cai, Minghang Zheng, Xin Jin, Yang LiuACM MM 2025 · 被引用 1 次
