SALAD: Part-Level Latent Diffusion for 3D Shape Generation and Manipulation
Juil Koo, Seungwoo Yoo, Minh Hieu Nguyen, Minhyuk Sung
摘要
We present a cascaded diffusion model based on a part-level implicit 3D representation. Our model achieves state-of-the-art generation quality and also enables part-level shape editing and manipulation without any additional training in conditional setup. Diffusion models have demonstrated impressive capabilities in data generation as well as zero-shot completion and editing via a guided reverse process. Recent research on 3D diffusion models has focused on improving their generation capabilities with various data representations, while the absence of structural information has limited their capability in completion and editing tasks. We thus propose our novel diffusion model using a part-level implicit representation. To effectively learn diffusion with high-dimensional embedding vectors of parts, we propose a cascaded framework, learning diffusion first on a low-dimensional subspace encoding extrinsic parameters of parts and then on the other high-dimensional subspace encoding intrinsic attributes. In the experiments, we demonstrate the outperformance of our method compared with the previous ones both in generation and part-level completion and manipulation tasks. Our project page is https://salad3d.github.io.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper40
- PartCrafter: Structured 3D Mesh Generation via Compositional Latent Diffusion TransformersYuchen Lin, Chenguo Lin, Panwang Pan, Honglei Yan 等NeurIPS 2025 · 被引用 89 次
- HoloPart: Generative 3D Part Amodal SegmentationYunhan Yang, Yuanchen Guo, Yukun Huang, Zi-Xin Zou 等ICLR 2026 · 被引用 62 次
- AutoPartGen: Autoregressive 3D Part Generation and DiscoveryMinghao Chen, Jianyuan Wang, Roman Shapovalov, Tom Monnier 等NeurIPS 2025 · 被引用 29 次
- DiffCAD: Weakly-Supervised Probabilistic CAD Model Retrieval and Alignment from an RGB ImageDaoyi Gao, Dávid Rozenberszki, Stefan Leutenegger, Angela DaiSIGGRAPH 2024 · 被引用 28 次
- Part123: Part-aware 3D Reconstruction from a Single-view ImageAnran Liu, Cheng Lin, Yuan Liu, Xiaoxiao Long 等SIGGRAPH 2024 · 被引用 23 次
它引用的顶会 Paper34
- Learning Transferable Visual Models From Natural Language SupervisionAlec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh 等ICML 2021 · 被引用 47,906 次
- Denoising Diffusion Probabilistic ModelsJonathan Ho, Ajay Jain, Pieter AbbeelNeurIPS 2020 · 被引用 35,902 次
- Diffusion Models Beat GANs on Image SynthesisPrafulla Dhariwal, Alexander Quinn NicholNeurIPS 2021 · 被引用 13,211 次
- High-Resolution Image Synthesis with Latent Diffusion ModelsRobin Rombach, Andreas Blattmann, Dominik Lorenz, Patrick Esser 等CVPR 2022 · 被引用 13,123 次
- SDEdit: Guided Image Synthesis and Editing with Stochastic Differential EquationsChenlin Meng, Yutong He, Yang Song, Jiaming Song 等ICLR 2022 · 被引用 2,128 次
相关 Paper
- SPAGHETTI: editing implicit shapes through part aware generationAmir Hertz, Or Perel, Raja Giryes, Olga Sorkine-Hornung 等SIGGRAPH 2022 · 被引用 59 次
- Diffusion-SDF: Conditional Generative Modeling of Signed Distance FunctionsGene Chou, Yuval Bahat, Felix HeideICCV 2023 · 被引用 171 次
- 3D Semantic Subspace Traverser: Empowering 3D Generative Model with Shape Editing CapabilityRuowei Wang, Yu Liu, Pei Su, Jianwei Zhang 等ICCV 2023 · 被引用 1 次
- SALAD: Skeleton-aware Latent Diffusion for Text-driven Motion Generation and EditingSeokhyeon Hong, Chaelin Kim, Serin Yoon, Junghyun Nam 等CVPR 2025
- DiffFacto: Controllable Part-Based 3D Point Cloud Generation with Cross DiffusionGeorge Kiyohiro Nakayama, Mikaela Angelina Uy, Jiahui Huang, Shi-Min Hu 等ICCV 2023 · 被引用 46 次
