In-Hand 3D Object Scanning from an RGB Sequence
Shreyas Hampali, Tomas Hodan, Luan Tran, Lingni Ma, Cem Keskin, Vincent Lepetit
Abstract
We propose a method for in-hand 3D scanning of an unknown object with a monocular camera. Our method relies on a neural implicit surface representation that captures both the geometry and the appearance of the object, however, by contrast with most NeRF-based methods, we do not assume that the camera-object relative poses are known. Instead, we simultaneously optimize both the object shape and the pose trajectory. As direct optimization over all shape and pose parameters is prone to fail without coarselevel initialization, we propose an incremental approach that starts by splitting the sequence into carefully selected overlapping segments within which the optimization is likely to succeed. We reconstruct the object shape and track its poses independently within each segment, then merge all the segments before performing a global optimization. We show that our method is able to reconstruct the shape and color of both textured and challenging texture-less objects, outperforms classical methods that rely only on appearance features, and that its performance is close to recent methods that assume known camera poses.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers10
- Diffusion-Guided Reconstruction of Everyday Hand-Object Interaction ClipsYufei Ye, Poorvi Hebbar, Abhinav Gupta, Shubham TulsianiICCV 2023 · 80 citations
- G-HOP: Generative Hand-Object Prior for Interaction Reconstruction and Grasp SynthesisYufei Ye, Abhinav Gupta, Kris Kitani, Shubham TulsianiCVPR 2024 · 13 citations
- HORT: Monocular Hand-held Objects Reconstruction with TransformersZerui Chen, Rolandos Alexandros Potamias, Shizhe Chen, Cordelia SchmidICCV 2025 · 4 citations
- Fully Dynamic Algorithms for Chamfer DistanceGramoz Goranci, Shaofeng H.-C. Jiang, Peter Kiss, Eva Szilagyi et al.NeurIPS 2025 · 3 citations
- MagicHOI: Leveraging 3D Priors for Accurate Hand-Object Reconstruction from Short Monocular Video ClipsShibo Wang, Haonan He, Maria Parelli, Christoph Gebhardt et al.ICCV 2025 · 2 citations
Builds on11
- Multiview Neural Surface Reconstruction by Disentangling Geometry and AppearanceLior Yariv, Yoni Kasten, Dror Moran, Meirav Galun et al.NeurIPS 2020 · 1,010 citations
- UNISURF: Unifying Neural Implicit Surfaces and Radiance Fields for Multi-View ReconstructionMichael Oechsle, Songyou Peng, Andreas GeigerICCV 2021 · 885 citations
- BARF: Bundle-Adjusting Neural Radiance FieldsChen-Hsuan Lin, Wei-Chiu Ma, Antonio Torralba, Simon LuceyICCV 2021 · 867 citations
- Self-Calibrating Neural Radiance FieldsYoonwoo Jeong, Seokjun Ahn, Christopher B. Choy, Animashree Anandkumar et al.ICCV 2021 · 275 citations
- SAMURAI: Shape And Material from Unconstrained Real-world Arbitrary Image collectionsMark Boss, Andreas Engelhardt, Abhishek Kar, Yuanzhen Li et al.NeurIPS 2022 · 104 citations
Related papers
- Free-Moving Object Reconstruction and Pose Estimation with Virtual CameraHaixin Shi, Yinlin Hu, Daniel Koguciuk, Juan-Ting Lin et al.AAAI 2025 · 2 citations
- Neural Surface Reconstruction of Dynamic Scenes with Monocular RGB-D CameraHongrui Cai, Wanquan Feng, Xuetao Feng, Yan Wang et al.NeurIPS 2022 · 83 citations
- GNeRF: GAN-based Neural Radiance Field without Posed CameraQuan Meng, Anpei Chen, Haimin Luo, Minye Wu et al.ICCV 2021 · 222 citations
- NoPe-NeRF: Optimising Neural Radiance Field with No Pose PriorWenjing Bian, Zirui Wang, Kejie Li, Jia-Wang BianCVPR 2023
- CodeNeRF: Disentangled Neural Radiance Fields for Object CategoriesWonbong Jang, Lourdes AgapitoICCV 2021 · 246 citations
