DAE-Net: Deforming Auto-Encoder for fine-grained shape co-segmentation
Zhiqin Chen, Qimin Chen, Hang Zhou, Hao Zhang
摘要
We present an unsupervised 3D shape co-segmentation method which learns a set of deformable part templates from a shape collection. To accommodate structural variations in the collection, our network composes each shape by a selected subset of template parts which are affine-transformed. To maximize the expressive power of the part templates, we introduce a per-part deformation network to enable the modeling of diverse parts with substantial geometry variations, while imposing constraints on the deformation capacity to ensure fidelity to the originally represented parts. We also propose a training scheme to effectively overcome local minima. Architecturally, our network is a branched autoencoder, with a CNN encoder taking a voxel shape as input and producing per-part transformation matrices, latent codes, and part existence scores, and the decoder outputting point occupancies to define the reconstruction loss. Our network, coined DAE-Net for Deforming Auto-Encoder, can achieve unsupervised 3D shape co-segmentation that yields fine-grained, compact, and meaningful parts that are consistent across diverse shapes. We conduct extensive experiments on the ShapeNet Part dataset, DFAUST, and an animal subset of Objaverse to show superior performance over prior methods. Code and data are available at https://github.com/czq142857/DAE-Net.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper2
- GenAnalysis: Joint Shape Analysis by Learning Man-Made Shape Generators with Deformation RegularizationsYuezhi Yang, Haitao Yang, Kiyohiro Nakayama, Xiangru Huang 等SIGGRAPH 2025 · 被引用 2 次
- Self-Supervised Learning of Hybrid Part-Aware 3D Representations of 2D Gaussians and SuperquadricsZhirui Gao, Renjiao Yi, Yuhang Huang, Wei Chen 等ICCV 2025 · 被引用 2 次
它引用的顶会 Paper32
- Learning Transferable Visual Models From Natural Language SupervisionAlec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh 等ICML 2021 · 被引用 47,906 次
- Photorealistic Text-to-Image Diffusion Models with Deep Language UnderstandingChitwan Saharia, William Chan, Saurabh Saxena, Lala Li 等NeurIPS 2022 · 被引用 8,965 次
- Emerging Properties in Self-Supervised Vision TransformersMathilde Caron, Hugo Touvron, Ishan Misra, Hervé Jégou 等ICCV 2021 · 被引用 8,921 次
- Flamingo: a Visual Language Model for Few-Shot LearningJean-Baptiste Alayrac, Jeff Donahue, Pauline Luc, Antoine Miech 等NeurIPS 2022 · 被引用 6,707 次
- Scaling Up Visual and Vision-Language Representation Learning With Noisy Text SupervisionChao Jia, Yinfei Yang, Ye Xia, Yi-Ting Chen 等ICML 2021 · 被引用 5,401 次
相关 Paper
- BAE-NET: Branched Autoencoder for Shape Co-SegmentationZhiqin Chen, Kangxue Yin, Matthew Fisher, Siddhartha Chaudhuri 等ICCV 2019 · 被引用 153 次
- EditVAE: Unsupervised Parts-Aware Controllable 3D Point Cloud Shape GenerationShidi Li, Miaomiao Liu, Christian WalderAAAI 2022 · 被引用 35 次
- Learning Part Generation and Assembly for Structure-Aware Shape SynthesisJun Li, Chengjie Niu, Kai XuAAAI 2020 · 被引用 85 次
- Composite Shape Modeling via Latent Space FactorizationAnastasia Dubrovina, Fei Xia, Panos Achlioptas, Mira Shalah 等ICCV 2019 · 被引用 66 次
- Few-Shot Learning of Part-Specific Probability Space for 3D Shape SegmentationLingjing Wang, Xiang Li, Yi FangCVPR 2020
