GECCO: Geometrically-Conditioned Point Diffusion Models
Michal J. Tyszkiewicz, Pascal Fua, Eduard Trulls
Abstract
Diffusion models generating images conditionally on text, such as Dall-E 2 [51] and Stable Diffusion [53] , have recently made a splash far beyond the computer vision community. Here, we tackle the related problem of generating point clouds, both unconditionally, and conditionally with images. For the latter, we introduce a novel geometricallymotivated conditioning scheme based on projecting sparse image features into the point cloud and attaching them to each individual point, at every step in the denoising process. This approach improves geometric consistency and yields greater fidelity than current methods relying on unstructured, global latent codes. Additionally, we show how to apply recent continuous-time diffusion schemes [59, 21] . Our method performs on par or above the state of art on conditional and unconditional experiments on synthetic data, while being faster, lighter, and delivering tractable likelihoods. We show it can also scale to diverse indoors scenes. … ; INPUT NOISE (XYZ) CONDITIONING FEATURES (t1) ; CNN DDM (t1) DDM (t2) … DDM (tN) POINT CLOUD (t2) CONDITIONING FEATURES (t2)
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers9
- Towards Realistic Scene Generation with LiDAR Diffusion ModelsHaoxi Ran, Vitor Guizilini, Yue WangCVPR 2024 · 26 citations
- Reconstruction of Manipulated Garment with Guided Deformation PriorRen Li, Corentin Dumery, Zhantao Deng, Pascal FuaNeurIPS 2024 · 10 citations
- Template Free Reconstruction of Human-object Interaction with Procedural Interaction GenerationXianghui Xie, Bharat Lal Bhatnagar, Jan Eric Lenssen, Gerard Pons-MollCVPR 2024 · 6 citations
- Single View Garment Reconstruction Using Diffusion Mapping Via Pattern CoordinatesRen Li, Cong Cao, Corentin Dumery, Yingxuan You et al.SIGGRAPH 2025 · 5 citations
- MamTiff-CAD: Multi-Scale Latent Diffusion with Mamba+ for Complex Parametric SequenceLiyuan Deng, Yunpeng Bai, Yongkang Dai, Xiaoshui Huang et al.ICCV 2025 · 3 citations
Builds on34
- Learning Transferable Visual Models From Natural Language SupervisionAlec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh et al.ICML 2021 · 47,906 citations
- Denoising Diffusion Probabilistic ModelsJonathan Ho, Ajay Jain, Pieter AbbeelNeurIPS 2020 · 35,902 citations
- High-Resolution Image Synthesis with Latent Diffusion ModelsRobin Rombach, Andreas Blattmann, Dominik Lorenz, Patrick Esser et al.CVPR 2022 · 13,123 citations
- Photorealistic Text-to-Image Diffusion Models with Deep Language UnderstandingChitwan Saharia, William Chan, Saurabh Saxena, Lala Li et al.NeurIPS 2022 · 8,965 citations
- A ConvNet for the 2020sZhuang Liu, Hanzi Mao, Chao-Yuan Wu, Christoph Feichtenhofer et al.CVPR 2022 · 6,782 citations
Related papers
- PC2: Projection-Conditioned Point Cloud Diffusion for Single-Image 3D ReconstructionLuke Melas-Kyriazi, Christian Rupprecht, Andrea VedaldiCVPR 2023
- Harnessing Text-to-Image Diffusion Models for Point Cloud Self-Supervised LearningYiyang Chen, Shanshan Zhao, Lunhao Duan, Changxing Ding et al.ICCV 2025
- Sketch and Text Guided Diffusion Model for Colored Point Cloud GenerationZijie Wu, Yaonan Wang, Mingtao Feng, He Xie et al.ICCV 2023 · 55 citations
- Points-to-3D: Bridging the Gap between Sparse Points and Shape-Controllable Text-to-3D GenerationChaohui Yu, Qiang Zhou, Jingliang Li, Zhe Zhang et al.ACM MM 2023 · 26 citations
- Points-to-3D: Structure-Aware 3D Generation with Point Cloud PriorsJiatong Xia, Zicheng Duan, Anton van den Hengel, Lingqiao LiuCVPR 2026 · 6 citations
