Atlas Gaussians Diffusion for 3D Generation
Haitao Yang, Yuan Dong, Hanwen Jiang, Dejia Xu, Georgios Pavlakos, Qixing Huang
摘要
Using the latent diffusion model has proven effective in developing novel 3D generation techniques. To harness the latent diffusion model, a key challenge is designing a high-fidelity and efficient representation that links the latent space and the 3D space. In this paper, we introduce Atlas Gaussians, a novel representation for feed-forward native 3D generation. Atlas Gaussians represent a shape as the union of local patches, and each patch can decode 3D Gaussians. We parameterize a patch as a sequence of feature vectors and design a learnable function to decode 3D Gaussians from the feature vectors. In this process, we incorporate UV-based sampling, enabling the generation of a sufficiently large, and theoretically infinite, number of 3D Gaussian points. The large amount of 3D Gaussians enables the generation of high-quality details. Moreover, due to local awareness of the representation, the transformer-based decoding procedure operates on a patch level, ensuring efficiency. We train a variational autoencoder to learn the Atlas Gaussians representation, and then apply a latent diffusion model on its latent space for learning 3D Generation. Experiments show that our approach outperforms the prior arts of feed-forward native 3D generation. Project page: https://yanghtr.github.io/projects/atlas_gaussians.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper6
- Native and Compact Structured Latents for 3D GenerationJianfeng Xiang, Xiaoxue Chen, Sicheng Xu, Ruicheng Wang 等CVPR 2026 · 被引用 177 次
- SIGMAN: Scaling 3D Human Gaussian Generation with Millions of AssetsYuhang Yang, Fengqi Liu, Yixing Lu, Qin Zhao 等ICCV 2025 · 被引用 5 次
- Gaussian Variation Field Diffusion for High-Fidelity Video-to-4D SynthesisBowen Zhang, Sicheng Xu, Chuxin Wang, Jiaolong Yang 等ICCV 2025 · 被引用 4 次
- Text-Image Conditioned 3D GenerationJiazhong Cen, Jiemin Fang, Sikuang Li, Guanjun Wu 等CVPR 2026 · 被引用 1 次
- Generative 3D Gaussians with Learned Density ControlRunjie Yan, Yan-Pei Cao, Peng Wang, Ding Liang 等SIGGRAPH 2026
它引用的顶会 Paper59
- Learning Transferable Visual Models From Natural Language SupervisionAlec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh 等ICML 2021 · 被引用 47,906 次
- Denoising Diffusion Probabilistic ModelsJonathan Ho, Ajay Jain, Pieter AbbeelNeurIPS 2020 · 被引用 35,902 次
- An Image is Worth 16x16 Words: Transformers for Image Recognition at ScaleAlexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn 等ICLR 2021 · 被引用 21,477 次
- High-Resolution Image Synthesis with Latent Diffusion ModelsRobin Rombach, Andreas Blattmann, Dominik Lorenz, Patrick Esser 等CVPR 2022 · 被引用 13,123 次
- Denoising Diffusion Implicit ModelsJiaming Song, Chenlin Meng, Stefano ErmonICLR 2021 · 被引用 11,743 次
相关 Paper
- Repurposing 2D Diffusion Models with Gaussian Atlas for 3D GenerationTiange Xiang, Kai Li, Chengjiang Long, Christian Häne 等ICCV 2025 · 被引用 1 次
- DiffGS: Functional Gaussian Splatting DiffusionJunsheng Zhou, Weiqi Zhang, Yu-Shen LiuNeurIPS 2024 · 被引用 69 次
- Prometheus: 3D-Aware Latent Diffusion Models for Feed-Forward Text-to-3D Scene GenerationYuanbo Yang, Jiahao Shao, Xinyang Li, Yujun Shen 等CVPR 2025
- E3Gen: Efficient, Expressive and Editable Avatars GenerationWeitian Zhang, Yichao Yan, Yunhui Liu, Xingdong Sheng 等ACM MM 2024 · 被引用 4 次
- DirectTriGS: Triplane-based Gaussian Splatting Field Representation for 3D GenerationXiaoliang Ju, Hongsheng LiCVPR 2025
