Large Point-to-Gaussian Model for Image-to-3D Generation
Longfei Lu, Huachen Gao, Tao Dai, Yaohua Zha, Zhi Hou, Junta Wu, Shu-Tao Xia
摘要
Recently, image-to-3D approaches have significantly advanced the generation quality and speed of 3D assets based on large reconstruction models, particularly 3D Gaussian reconstruction models. Existing large 3D Gaussian models directly map 2D image to 3D Gaussian parameters, while regressing 2D image to 3D Gaussian representations is challenging without 3D priors. In this paper, we propose a large Point-to-Gaussian model, that inputs the initial point cloud produced from large 3D diffusion model conditional on 2D image to generate the Gaussian parameters, for image-to-3D generation. The point cloud provides initial 3D geometry prior for Gaussian generation, thus significantly facilitating image-to-3D Generation. Moreover, we present the Attention mechanism, Projection mechanism, and Point feature extractor, dubbed as APP block, for fusing the image features with point cloud features. The qualitative and quantitative experiments extensively demonstrate the effectiveness of the proposed approach on GSO and Objaverse datasets, and show the proposed method achieves state-of-the-art performance.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper7
- RemVerse: Supporting Reminiscence Activities for Older Adults through AI-Assisted Virtual RealityRuohao Li, Jiawei Li, Jia Sun, Zhiqing Wu 等UbiComp 2025 · 被引用 8 次
- GaussianGrow: Geometry-aware Gaussian Growing from 3D Point Clouds with Text GuidanceWeiqi Zhang, Junsheng Zhou, Haotian Geng, Kanle Shi 等CVPR 2026 · 被引用 2 次
- GAP: Gaussianize Any Point Clouds with Text GuidanceWeiqi Zhang, Junsheng Zhou, Haotian Geng, Wenyuan Zhang 等ICCV 2025 · 被引用 2 次
- You See it, You Got it: Learning 3D Creation on Pose-Free Videos at ScaleBaorui Ma, Huachen Gao, Haoge Deng, Zhengxiong Luo 等CVPR 2025
- Dehallu3D: Hallucination-Mitigated 3D Generation from a Single Image via Cyclic View Consistency RefinementXiwen Wang, Shichao Zhang, Ruowei Wang, Mao Li 等CVPR 2026
它引用的顶会 Paper40
- High-Resolution Image Synthesis with Latent Diffusion ModelsRobin Rombach, Andreas Blattmann, Dominik Lorenz, Patrick Esser 等CVPR 2022 · 被引用 13,123 次
- Zero-1-to-3: Zero-shot One Image to 3D ObjectRuoshi Liu, Rundi Wu, Basile Van Hoorick, Pavel Tokmakov 等ICCV 2023 · 被引用 1,662 次
- ProlificDreamer: High-Fidelity and Diverse Text-to-3D Generation with Variational Score DistillationZhengyi Wang, Cheng Lu, Yikai Wang, Fan Bao 等NeurIPS 2023 · 被引用 1,498 次
- MVDream: Multi-view Diffusion for 3D GenerationYichun Shi, Peng Wang, Jianglong Ye, Long Mai 等ICLR 2024 · 被引用 973 次
- DreamGaussian: Generative Gaussian Splatting for Efficient 3D Content CreationJiaxiang Tang, Jiawei Ren, Hang Zhou, Ziwei Liu 等ICLR 2024 · 被引用 955 次
相关 Paper
- Points-to-3D: Structure-Aware 3D Generation with Point Cloud PriorsJiatong Xia, Zicheng Duan, Anton van den Hengel, Lingqiao LiuCVPR 2026 · 被引用 6 次
- Repurposing 2D Diffusion Models with Gaussian Atlas for 3D GenerationTiange Xiang, Kai Li, Chengjiang Long, Christian Häne 等ICCV 2025 · 被引用 1 次
- GaussianDreamer: Fast Generation from Text to 3D Gaussians by Bridging 2D and 3D Diffusion ModelsTaoran Yi, Jiemin Fang, Junjie Wang, Guanjun Wu 等CVPR 2024 · 被引用 106 次
- Atlas Gaussians Diffusion for 3D GenerationHaitao Yang, Yuan Dong, Hanwen Jiang, Dejia Xu 等ICLR 2025
- Prometheus: 3D-Aware Latent Diffusion Models for Feed-Forward Text-to-3D Scene GenerationYuanbo Yang, Jiahao Shao, Xinyang Li, Yujun Shen 等CVPR 2025
