VV-Net: Voxel VAE Net With Group Convolutions for Point Cloud Segmentation
Hsien-Yu Meng, Lin Gao, Yu-Kun Lai, Dinesh Manocha
Abstract
We present a novel algorithm for point cloud segmentation.Our approach transforms unstructured point clouds into regular voxel grids, and further uses a kernel-based interpolated variational autoencoder (VAE) architecture to encode the local geometry within each voxel.Traditionally, the voxel representation only comprises Boolean occupancy information, which fails to capture the sparsely distributed points within voxels in a compact manner. In order to handle sparse distributions of points, we further employ radial basis functions (RBF) to compute a local, continuous representation within each voxel. Our approach results in a good volumetric representation that effectively tackles noisy point cloud datasets and is more robust for learning. Moreover, we further introduce group equivariant CNN to 3D, by defining the convolution operator on a symmetry group acting on and its isomorphic sets. This improves the expressive capacity without increasing parameters, leading to more robust segmentation results.We highlight the performance on standard benchmarks and show that our approach outperforms state-of-the-art segmentation algorithms on the ShapeNet and S3DIS datasets.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers28
- ePointDA: An End-to-End Simulation-to-Real Domain Adaptation Framework for LiDAR Point Cloud SegmentationSicheng Zhao, Yezhen Wang, Bo Li, Bichen Wu et al.AAAI 2021 · 112 citations
- Clustering based Point Cloud Representation Learning for 3D AnalysisTuo Feng, Wenguan Wang, Xiaohan Wang, Yi Yang et al.ICCV 2023 · 53 citations
- Reconstructing Interacting Hands with Interaction Prior from Monocular ImagesBinghui Zuo, Zimeng Zhao, Wenqian Sun, Wei Xie et al.ICCV 2023 · 26 citations
- RepKPU: Point Cloud Upsampling with Kernel Point Representation and DeformationYi Rong, Haoran Zhou, Kang Xia, Cheng Mei et al.CVPR 2024 · 22 citations
- Annotator: A Generic Active Learning Baseline for LiDAR Semantic SegmentationBinhui Xie, Shuang Li, Qingju Guo, Chi Harold Liu et al.NeurIPS 2023 · 21 citations
Related papers
- Interpolated Convolutional Networks for 3D Point Cloud UnderstandingJiageng Mao, Xiaogang Wang, Hongsheng LiICCV 2019 · 241 citations
- SE(3)-Equivariant Attention Networks for Shape Reconstruction in Function SpaceEvangelos Chatzipantazis, Stefanos Pertigkiozoglou, Edgar Dobriban, Kostas DaniilidisICLR 2023 · 7 citations
- Learning Coordinate-based Convolutional Kernels for Continuous SE(3) Equivariant and Efficient Point Cloud AnalysisJaein Kim, Hee Bin Yoo, Dong-Sig Han, Byoung-Tak ZhangCVPR 2026
- Rotation-Invariant Local-to-Global Representation Learning for 3D Point CloudSeohyun Kim, Jaeyoo Park, Bohyung HanNeurIPS 2020 · 92 citations
- EditVAE: Unsupervised Parts-Aware Controllable 3D Point Cloud Shape GenerationShidi Li, Miaomiao Liu, Christian WalderAAAI 2022 · 35 citations
