ZeroGrasp: Zero-Shot Shape Reconstruction Enabled Robotic Grasping
Shun Iwase, Muhammad Zubair Irshad, Katherine Liu, Vitor Guizilini, Robert Lee, Takuya Ikeda, Ayako Amma, Koichi Nishiwaki, Kris Kitani, Rares Ambrus, Sergey Zakharov
Abstract
Robotic grasping is a cornerstone capability of embodied systems. Many methods directly output grasps from partial information without modeling the geometry of the scene, leading to suboptimal motion and even collisions. To address these issues, we introduce ZeroGrasp, a novel framework that simultaneously performs 3D reconstruction and grasp pose prediction in near real-time. A key insight of our method is that occlusion reasoning and modeling the spatial relationships between objects is beneficial for both accurate reconstruction and grasping. We couple our method with a novel large-scale synthetic dataset, which comprises 1M photo-realistic images, high-resolution 3D reconstructions and 11.3B physically-valid grasp pose annotations for 12K objects from the Objaverse-LVIS dataset. We evaluate ZeroGrasp on the GraspNet-1B benchmark as well as through real-world robot experiments. Zero-Grasp achieves state-of-the-art performance and generalizes to novel real-world objects by leveraging synthetic data.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 57b71c90-0df6-4f75-a15c-8360d42bd3c9Cited by top-tier papers4
- LaS-Comp: Zero-shot 3D Completion with Latent–Spatial ConsistencyWeilong Yan, Li Haipeng, Hao Xu, Nianjin Ye et al.CVPR 2026 · 14 citations
- RealVLG-R1: A Large-Scale Real-World Visual-Language Grounding Benchmark for Robotic Perception and ManipulationLinfei Li, Lin Zhang, Ying ShenCVPR 2026
- TOSC: Task-Oriented Shape Completion for Open-World Dexterous Grasp Generation from Partial Point CloudsWeishang Wu, Yifei Shi, Zhiping CaiAAAI 2026
- A Cross-view Fusion Framework for Robust 6-DoF Grasp Pose EstimationKangjian Zhu, Haobo Jiang, Jianjun Qian, Jin XieCVPR 2026
Builds on29
- Zero-1-to-3: Zero-shot One Image to 3D ObjectRuoshi Liu, Rundi Wu, Basile Van Hoorick, Pavel Tokmakov et al.ICCV 2023 · 1,662 citations
- 6-DOF GraspNet: Variational Grasp Generation for Object ManipulationArsalan Mousavian, Clemens Eppner, Dieter FoxICCV 2019 · 673 citations
- Deep Marching Tetrahedra: a Hybrid Representation for High-Resolution 3D Shape SynthesisTianchang Shen, Jun Gao, Kangxue Yin, Ming-Yu Liu et al.NeurIPS 2021 · 652 citations
- PoinTr: Diverse Point Cloud Completion with Geometry-Aware TransformersXumin Yu, Yongming Rao, Ziyi Wang, Zuyan Liu et al.ICCV 2021 · 592 citations
- HuMoR: 3D Human Motion Model for Robust Pose EstimationDavis Rempe, Tolga Birdal, Aaron Hertzmann, Jimei Yang et al.ICCV 2021 · 398 citations
Related papers
- LRM-Zero: Training Large Reconstruction Models with Synthesized DataDesai Xie, Sai Bi, Zhixin Shu, Kai Zhang et al.NeurIPS 2024 · 36 citations
- GraphGrasp: Lightweight and Efficient Graph-Guided 6-DoF Robotic Grasp Pose Estimation NetworkSheng Yu, Di-Hua Zhai, Yuanqing XiaAAAI 2026
- DexVLG: Dexterous Vision-Language-Grasp Model at ScaleJiawei He, Danshi Li, Xinqiang Yu, Zekun Qi et al.ICCV 2025 · 6 citations
- Robust 3D Shape Reconstruction in Zero-Shot from a Single Image in the WildJunhyeong Cho, Kim Youwang, Hunmin Yang, Tae-Hyun OhCVPR 2025
- FLEX: Full-Body Grasping Without Full-Body GraspsPurva Tendulkar, Dídac Surís, Carl VondrickCVPR 2023
