SPAR3D: Stable Point-Aware Reconstruction of 3D Objects from Single Images
Zixuan Huang, Mark Boss, Aaryaman Vasishta, James M. Rehg, Varun Jampani
Abstract
We study the problem of single-image 3D object reconstruction. Recent works have diverged into two directions: regression-based modeling and generative modeling. Regression methods efficiently infer visible surfaces, but struggle with occluded regions. Generative methods handle uncertain regions better by modeling distributions, but are computationally expensive and the generation is often misaligned with visible surfaces. In this paper, we present SPAR3D, a novel two-stage approach aiming to take the best of both directions. The first stage of SPAR3D generates sparse 3D point clouds using a lightweight point diffusion model, which has a fast sampling speed. The second stage uses both the sampled point cloud and the input image to create highly detailed meshes. Our two-stage design enables probabilistic modeling of the ill-posed single-image 3D task while maintaining high computational efficiency and great output fidelity. Using point clouds as an intermediate representation further allows for interactive user edits. Evaluated on diverse datasets, SPAR3D demonstrates superior performance over previous state-of-the-art methods, at an inference speed of 0.7 seconds. Project page with code and model: https://spar3d.github.io
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 0e4c4797-c008-4d57-991d-18333822745aCited by top-tier papers21
- ShapeLLM-Omni: A Native Multimodal LLM for 3D Generation and UnderstandingJunliang Ye, Zhengyi Wang, Ruowen Zhao, Shenghao Xie et al.NeurIPS 2025 · 42 citations
- DeepMesh: Auto-Regressive Artist-Mesh Creation with Reinforcement LearningRuowen Zhao, Junliang Ye, Zhengyi Wang, Guangce Liu et al.ICCV 2025 · 9 citations
- V2M4: 4D Mesh Animation Reconstruction from a Single Monocular VideoJianqi Chen, Biao Zhang, Xiangjun Tang, Peter WonkaICCV 2025 · 8 citations
- Points-to-3D: Structure-Aware 3D Generation with Point Cloud PriorsJiatong Xia, Zicheng Duan, Anton van den Hengel, Lingqiao LiuCVPR 2026 · 6 citations
- Topology-Preserved Auto-regressive Mesh Generation in the Manner of Weaving SilkGaochao Song, Zibo Zhao, Haohan Weng, Jingbo Zeng et al.ICLR 2026 · 6 citations
Builds on34
- Denoising Diffusion Probabilistic ModelsJonathan Ho, Ajay Jain, Pieter AbbeelNeurIPS 2020 · 35,902 citations
- High-Resolution Image Synthesis with Latent Diffusion ModelsRobin Rombach, Andreas Blattmann, Dominik Lorenz, Patrick Esser et al.CVPR 2022 · 13,123 citations
- Denoising Diffusion Implicit ModelsJiaming Song, Chenlin Meng, Stefano ErmonICLR 2021 · 11,743 citations
- Zero-1-to-3: Zero-shot One Image to 3D ObjectRuoshi Liu, Rundi Wu, Basile Van Hoorick, Pavel Tokmakov et al.ICCV 2023 · 1,662 citations
- LRM: Large Reconstruction Model for Single Image to 3DYicong Hong, Kai Zhang, Jiuxiang Gu, Sai Bi et al.ICLR 2024 · 813 citations
Related papers
- PC2: Projection-Conditioned Point Cloud Diffusion for Single-Image 3D ReconstructionLuke Melas-Kyriazi, Christian Rupprecht, Andrea VedaldiCVPR 2023
- Make-It-3D: High-Fidelity 3D Creation from A Single Image with Diffusion PriorJunshu Tang, Tengfei Wang, Bo Zhang, Ting Zhang et al.ICCV 2023 · 405 citations
- One-2-3-45: Any Single Image to 3D Mesh in 45 Seconds without Per-Shape OptimizationMinghua Liu, Chao Xu, Haian Jin, Linghao Chen et al.NeurIPS 2023 · 755 citations
- Triplane Meets Gaussian Splatting: Fast and Generalizable Single-View 3D Reconstruction with TransformersZi-Xin Zou, Zhipeng Yu, Yuan-Chen Guo, Yangguang Li et al.CVPR 2024 · 119 citations
- Controllable Mesh Generation Through Sparse Latent Point Diffusion ModelsZhaoyang Lyu, Jinyi Wang, Yuwei An, Ya Zhang et al.CVPR 2023
