MoreFusion: Multi-object Reasoning for 6D Pose Estimation from Volumetric Fusion
Kentaro Wada, Edgar Sucar, Stephen James, Daniel Lenton, Andrew J. Davison
Abstract
Robots and other smart devices need efficient objectbased scene representations from their on-board vision systems to reason about contact, physics and occlusion. Recognized precise object models will play an important role alongside non-parametric reconstructions of unrecognized structures. We present a system which can estimate the accurate poses of multiple known objects in contact and occlusion from real-time, embodied multi-view vision. Our approach makes 3D object pose proposals from single RGB-D views, accumulates pose estimates and non-parametric occupancy information from multiple views as the camera moves, and performs joint optimization to estimate consistent, non-intersecting poses for multiple objects in contact.
We verify the accuracy and robustness of our approach experimentally on 2 object datasets: YCB-Video, and our own challenging Cluttered YCB-Video. We demonstrate a real-time robotics application where a robot arm precisely and orderly disassembles complicated piles of objects, using only on-board RGB-D vision.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 0e6fbd6b-88b5-4a79-a617-8ae9d742c40dCited by top-tier papers14
- SAR-Net: Shape Alignment and Recovery Network for Category-level 6D Object Pose and Size EstimationHaitao Lin, Zichang Liu, Chilam Cheang, Yanwei Fu et al.CVPR 2022 · 86 citations
- RNNPose: Recurrent 6-DoF Object Pose Refinement with Robust Correspondence Field Estimation and Pose OptimizationYan Xu, Kwan-Yee Lin, Guofeng Zhang, Xiaogang Wang et al.CVPR 2022 · 82 citations
- Coarse-to-Fine Q-attention: Efficient Learning for Visual Robotic Manipulation via DiscretisationStephen James, Kentaro Wada, Tristan Laidlow, Andrew J. DavisonCVPR 2022 · 65 citations
- Uni6D: A Unified CNN Framework without Projection Breakdown for 6D Pose EstimationXiaoke Jiang, Donghai Li, Hao Chen, Ye Zheng et al.CVPR 2022 · 54 citations
- PR-GCN: A Deep Graph Convolutional Network with Point Refinement for 6D Pose EstimationGuangyuan Zhou, Huiqun Wang, Jiaxin Chen, Di HuangICCV 2021 · 45 citations
Related papers
- QuickPose: Real-time Multi-view Multi-person Pose Estimation in Crowded ScenesZhize Zhou, Qing Shuai, Yize Wang, Qi Fang et al.SIGGRAPH 2022 · 15 citations
- 3DP3: 3D Scene Perception via Probabilistic ProgrammingNishad Gothoskar, Marco F. Cusumano-Towner, Ben Zinberg, Matin Ghavamizadeh et al.NeurIPS 2021 · 59 citations
- Online Unsupervised Learning of the 3D Kinematic Structure of Arbitrary Rigid BodiesUrbano Miguel Nunes, Yiannis DemirisICCV 2019 · 4 citations
- EM-Fusion: Dynamic Object-Level SLAM With Probabilistic Data AssociationMichael Strecke, Jörg StücklerICCV 2019 · 86 citations
- 3D Neural Embedding Likelihood: Probabilistic Inverse Graphics for Robust 6D Pose EstimationGuangyao Zhou, Nishad Gothoskar, Lirui Wang, Joshua B. Tenenbaum et al.ICCV 2023 · 5 citations
