GraspGen-X: Cross-Embodiment 6-DOF Diffusion-based Grasping
Beining Han, Yu-Wei Chao, Erwin Coumans, Clemens Eppner, Jia Deng, Stan Birchfield, Adithyavairavan Murali
Abstract
We study cross-embodiment 6-DOF robot grasping. Unlike prior works, we require the model not only to generalize to novel objects / scenes but also to novel gripper morphologies and physical grasping processes. Our method extends diffusion model based generative 6-DOF grasping models to condition on the additional gripper's representation. We propose a swept-volume heuristic for encoding the gripper. We train our cross-embodiment model with procedural grippers and a large-scale dataset of 2 Billion grasps. In simulation experiments, our model has the best zero-shot generalization to novel real-world grippers and objects over baseline methods. Our model also serves as a good initialization for finetuning to adapt to novel grippers. In ablations, we demonstrate the efficiency of our sweep-volume gripper representation and our procedural gripper training dataset. Last, we show zero-shot generalization to real-world novel grippers for 6-DOF grasping, surpassing baselines in cross-embodiment generalization.
Figure 1: We introduce GraspGen-X, a cross-embodiment 6 DOF Grasping model trained with a large scale dataset of procedural grippers and 2 Billion simulated grasps. We achieve zero-shot generalization to both unknown objects as well as grippers in the real world.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Builds on12
- 6-DOF GraspNet: Variational Grasp Generation for Object ManipulationArsalan Mousavian, Clemens Eppner, Dieter FoxICCV 2019 · 673 citations
- One Policy to Control Them All: Shared Modular Policies for Agent-Agnostic ControlWenlong Huang, Igor Mordatch, Deepak PathakICML 2020 · 214 citations
- Approximate convex decomposition for 3D meshes with collision-aware concavity and tree searchXinyue Wei, Minghua Liu, Zhan Ling, Hao SuSIGGRAPH 2022 · 79 citations
- CaP-X: A Framework for Benchmarking and Improving Coding Agents for Robot ManipulationLetian Fu, Justin Yu, Karim El-Refai, Ethan Kou et al.ICML 2026 · 36 citations
- SpaceTools: Tool-Augmented Spatial Reasoning via Double Interactive RLSiyi Chen, Mikaela Angelina Uy, Chan Hee Song, Faisal Ladhak et al.CVPR 2026 · 24 citations
Related papers
- DexGrasp Anything: Towards Universal Robotic Dexterous Grasping with Physics AwarenessYiming Zhong, Qi Jiang, Jingyi Yu, Yuexin MaCVPR 2025
- Scalable and General Whole-Body Control for Cross-Humanoid LocomotionYufei Xue, Yunfeng Lin, Wentao Dong, Yang Tang et al.ICML 2026
- Cross-Embodiment Dexterous Grasping with Reinforcement LearningHaoqi Yuan, Bohan Zhou, Yuhui Fu, Zongqing LuICLR 2025
- TraceGen: World Modeling in 3D Trace Space Enables Learning from Cross-Embodiment VideosSeungjae Lee, Yoonkyo Jung, Inkook Chun, Yao-Chih Lee et al.CVPR 2026 · 17 citations
- GraphGrasp: Lightweight and Efficient Graph-Guided 6-DoF Robotic Grasp Pose Estimation NetworkSheng Yu, Di-Hua Zhai, Yuanqing XiaAAAI 2026
