Single-View 3D Object Reconstruction From Shape Priors in Memory
Shuo Yang, Min Xu, Haozhe Xie, Stuart W. Perry, Jiahao Xia
Abstract
Existing methods for single-view 3D object reconstruction directly learn to transform image features into 3D representations. However, these methods are vulnerable to images containing noisy backgrounds and heavy occlusions because the extracted image features do not contain enough information to reconstruct high-quality 3D shapes. Humans routinely use incomplete or noisy visual cues from an image to retrieve similar 3D shapes from their memory and reconstruct the 3D shape of an object. Inspired by this, we propose a novel method, named Mem3D, that explicitly constructs shape priors to supplement the missing information in the image. Specifically, the shape priors are in the forms of "image-voxel" pairs in the memory network, which is stored by a well-designed writing strategy during training. We also propose a voxel triplet loss function that helps to retrieve the precise 3D shapes that are highly related to the input image from shape priors. The LSTM-based shape encoder is introduced to extract information from the retrieved 3D shapes, which are useful in recovering the 3D shape of an object that is heavily occluded or in complex environments. Experimental results demonstrate that Mem3D significantly improves reconstruction quality and performs favorably against state-of-the-art methods on the ShapeNet and Pix3D datasets.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers8
- CAFE: Learning to Condense Dataset by Aligning FeaturesKai Wang, Bo Zhao, Xiangyu Peng, Zheng Zhu et al.CVPR 2022 · 140 citations
- Objects in Semantic TopologyShuo Yang, Peize Sun, Yi Jiang, Xiaobo Xia et al.ICLR 2022 · 35 citations
- U-RED: Unsupervised 3D Shape Retrieval and Deformation for Partial Point CloudsYan Di, Chenyangguang Zhang, Ruida Zhang, Fabian Manhardt et al.ICCV 2023 · 15 citations
- UMIFormer: Mining the Correlations between Similar Tokens for Multi-View 3D ReconstructionZhenwei Zhu, Liying Yang, Ning Li, Chaohao Jiang et al.ICCV 2023 · 12 citations
- Coupled Reconstruction of Cortical Surfaces by Diffeomorphic Mesh DeformationHao Zheng, Hongming Li, Yong FanNeurIPS 2023 · 11 citations
Builds on4
- Free Lunch for Few-shot Learning: Distribution CalibrationShuo Yang, Lu Liu, Min XuICLR 2021 · 378 citations
- Pix2Vox: Context-Aware 3D Reconstruction From Single and Multi-View ImagesHaozhe Xie, Hongxun Yao, Xiaoshuai Sun, Shangchen Zhou et al.ICCV 2019 · 373 citations
- Domain-Adaptive Single-View 3D ReconstructionPedro O. Pinheiro, Negar Rostamzadeh, Sungjin AhnICCV 2019 · 27 citations
- FroDO: From Detections to 3D ObjectsMartin Rünz, Kejie Li, Meng Tang, Lingni Ma et al.CVPR 2020
Related papers
- CDPNet: Cross-Modal Dual Phases Network for Point Cloud CompletionZhenjiang Du, Jiale Dou, Zhitao Liu, Jiwei Wei et al.AAAI 2024 · 17 citations
- From Image Collections to Point Clouds With Self-Supervised Shape and Pose NetworksNavaneet K. L., Ansu Mathew, Shashank Kashyap, Wei-Chih Hung et al.CVPR 2020
- Generating Point Cloud from Single Image in The Few Shot ScenarioYu Lin, Jinghui Guo, Yang Gao, Yi-Fan Li et al.ACM MM 2021 · 6 citations
- Toward Realistic Single-View 3D Object Reconstruction with Unsupervised Learning from Multiple ImagesLong-Nhat Ho, Anh Tuan Tran, Quynh Phung, Minh HoaiICCV 2021 · 10 citations
- Single Image 3D Object Estimation with Primitive Graph NetworksQian He, Desen Zhou, Bo Wan, Xuming HeACM MM 2021 · 1 citation
