SetVAE: Learning Hierarchical Composition for Generative Modeling of Set-Structured Data
Jinwoo Kim, Jaehoon Yoo, Juho Lee, Seunghoon Hong
Abstract
Generative modeling of set-structured data, such as point clouds, requires reasoning over local and global structures at various scales. However, adopting multi-scale frameworks for ordinary sequential data to a set-structured data is nontrivial as it should be invariant to the permutation of its elements. In this paper, we propose SetVAE, a hierarchical variational autoencoder for sets. Motivated by recent progress in set encoding, we build SetVAE upon attentive modules that first partition the set and project the partition back to the original cardinality. Exploiting this module, our hierarchical VAE learns latent variables at multiple scales, capturing coarse-to-fine dependency of the set elements while achieving permutation invariance. We evaluate our model on point cloud generation task and achieve competitive performance to the prior arts with substantially smaller model capacity. We qualitatively demonstrate that our model generalizes to unseen set sizes and learns interesting subset relations without supervision. Our implementation is available at https://github.com/ jw9730/setvae.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 320e8428-017c-470c-aab0-2a759f5a3c7bCited by top-tier papers25
- LION: Latent Point Diffusion Models for 3D Shape GenerationXiaohui Zeng, Arash Vahdat, Francis Williams, Zan Gojcic et al.NeurIPS 2022 · 752 citations
- DiT-3D: Exploring Plain Diffusion Transformers for 3D Shape GenerationShentong Mo, Enze Xie, Ruihang Chu, Lanqing Hong et al.NeurIPS 2023 · 157 citations
- Top-N: Equivariant Set and Graph Generation without ExchangeabilityClément Vignac, Pascal FrossardICLR 2022 · 42 citations
- 2D-3D Interlaced Transformer for Point Cloud Segmentation with Scene-Level SupervisionCheng-Kun Yang, Min-Hung Chen, Yung-Yu Chuang, Yen-Yu LinICCV 2023 · 30 citations
- GECCO: Geometrically-Conditioned Point Diffusion ModelsMichal J. Tyszkiewicz, Pascal Fua, Eduard TrullsICCV 2023 · 28 citations
Builds on6
- Object-Centric Learning with Slot AttentionFrancesco Locatello, Dirk Weissenborn, Thomas Unterthiner, Aravindh Mahendran et al.NeurIPS 2020 · 1,275 citations
- NVAE: A Deep Hierarchical Variational AutoencoderArash Vahdat, Jan KautzNeurIPS 2020 · 1,141 citations
- PointFlow: 3D Point Cloud Generation With Continuous Normalizing FlowsGuandao Yang, Xun Huang, Zekun Hao, Ming-Yu Liu et al.ICCV 2019 · 794 citations
- FSPool: Learning Set Representations with Featurewise Sort PoolingYan Zhang, Jonathon S. Hare, Adam Prügel-BennettICLR 2020 · 92 citations
- Exchangeable Neural ODE for Set ModelingYang Li, Haidong Yi, Christopher M. Bender, Siyuan Shan et al.NeurIPS 2020 · 32 citations
Related papers
- SCHA-VAE: Hierarchical Context Aggregation for Few-Shot GenerationGiorgio Giannone, Ole WintherICML 2022 · 11 citations
- EditVAE: Unsupervised Parts-Aware Controllable 3D Point Cloud Shape GenerationShidi Li, Miaomiao Liu, Christian WalderAAAI 2022 · 35 citations
- Efficient Hierarchical Entropy Model for Learned Point Cloud CompressionRui Song, Chunyang Fu, Shan Liu, Ge LiCVPR 2023
- VV-Net: Voxel VAE Net With Group Convolutions for Point Cloud SegmentationHsien-Yu Meng, Lin Gao, Yu-Kun Lai, Dinesh ManochaICCV 2019 · 268 citations
- Voxel Set Transformer: A Set-to-Set Approach to 3D Object Detection from Point CloudsChenhang He, Ruihuang Li, Shuai Li, Lei ZhangCVPR 2022 · 217 citations
