Live Semantic 3D Perception for Immersive Augmented Reality
Lei Han, Tian Zheng, Yinheng Zhu, Lan Xu, Lu Fang
Abstract
Semantic understanding of 3D environments is critical for both the unmanned system and the human involved virtual/augmented reality (VR/AR) immersive experience. Spatially-sparse convolution, taking advantage of the intrinsic sparsity of 3D point cloud data, makes high resolution 3D convolutional neural networks tractable with state-of-the-art results on 3D semantic segmentation problems. However, the exhaustive computations limits the practical usage of semantic 3D perception for VR/AR applications in portable devices. In this paper, we identify that the efficiency bottleneck lies in the unorganized memory access of the sparse convolution steps, i.e., the points are stored independently based on a predefined dictionary, which is inefficient due to the limited memory bandwidth of parallel computing devices (GPU). With the insight that points are continuous as 2D surfaces in 3D space, a chunk-based sparse convolution scheme is proposed to reuse the neighboring points within each spatially organized chunk. An efficient multi-layer adaptive fusion module is further proposed for employing the spatial consistency cue of 3D data to further reduce the computational burden. Quantitative experiments on public datasets demonstrate that our approach works 11× faster than previous approaches with competitive accuracy. By implementing both semantic and geometric 3D reconstruction simultaneously on a portable tablet device, we demo a foundation platform for immersive AR applications.
Ask about this paper
Ask your agent about it.
Lune has read the top-tier papers around this one, so every answer names the papers it rests on.
Your agent calls
Lunesearch_papers
Free to start. No credit card required.
Terminal
Install the CLIlune papers get 8211e8e2-aaf9-41ef-a8a5-5b6ab849a61fCited by top-tier papers12
- When XR and AI Meet - A Scoping Review on Extended Reality and Artificial IntelligenceTeresa Hirzle, Florian Müller, Fiona Draxler, Martin Schmitz et al.CHI 2023 · 90 citations
- ScalAR: Authoring Semantically Adaptive Augmented Reality Experiences in Virtual RealityXun Qian, Fengming He, Xiyun Hu, Tianyi Wang et al.CHI 2022 · 69 citations
- Voxel-based 3D Detection and Reconstruction of Multiple Objects from a Single ImageFeng Liu, Xiaoming LiuNeurIPS 2021 · 43 citations
- Training an Open-Vocabulary Monocular 3D Detection Model without 3D DataRui Huang, Henry Zheng, Yan Wang, Zhuofan Xia et al.NeurIPS 2024 · 26 citations
- MineXR: Mining Personalized Extended Reality InterfacesHyunsung Cho, Yukang Yan, Kashyap Todi, Mark Parent et al.CHI 2024 · 22 citations
Related papers
- Interpolation-Aware Padding for 3D Sparse Convolutional Neural NetworksYu-Qi Yang, Peng-Shuai Wang, Yang LiuICCV 2021 · 4 citations
- Not All Neighbors Matter: Point Distribution-Aware Pruning for 3D Point CloudYejin Lee, Donghyun Lee, JungUk Hong, Jae W. Lee et al.AAAI 2023 · 7 citations
- Fusion-Aware Point Convolution for Online Semantic 3D Scene SegmentationJiazhao Zhang, Chenyang Zhu, Lintao Zheng, Kai XuCVPR 2020
- Interpolated Convolutional Networks for 3D Point Cloud UnderstandingJiageng Mao, Xiaogang Wang, Hongsheng LiICCV 2019 · 241 citations
- INS-Conv: Incremental Sparse Convolution for Online 3D SegmentationLeyao Liu, Tian Zheng, Yun-Jou Lin, Kai Ni et al.CVPR 2022 · 19 citations
