Live Semantic 3D Perception for Immersive Augmented Reality
Lei Han, Tian Zheng, Yinheng Zhu, Lan Xu, Lu Fang
摘要
Semantic understanding of 3D environments is critical for both the unmanned system and the human involved virtual/augmented reality (VR/AR) immersive experience. Spatially-sparse convolution, taking advantage of the intrinsic sparsity of 3D point cloud data, makes high resolution 3D convolutional neural networks tractable with state-of-the-art results on 3D semantic segmentation problems. However, the exhaustive computations limits the practical usage of semantic 3D perception for VR/AR applications in portable devices. In this paper, we identify that the efficiency bottleneck lies in the unorganized memory access of the sparse convolution steps, i.e., the points are stored independently based on a predefined dictionary, which is inefficient due to the limited memory bandwidth of parallel computing devices (GPU). With the insight that points are continuous as 2D surfaces in 3D space, a chunk-based sparse convolution scheme is proposed to reuse the neighboring points within each spatially organized chunk. An efficient multi-layer adaptive fusion module is further proposed for employing the spatial consistency cue of 3D data to further reduce the computational burden. Quantitative experiments on public datasets demonstrate that our approach works 11× faster than previous approaches with competitive accuracy. By implementing both semantic and geometric 3D reconstruction simultaneously on a portable tablet device, we demo a foundation platform for immersive AR applications.
问问这篇 Paper
问问你的智能体。
Lune 读过与它相关的顶会 Paper,每个回答都会注明依据哪几篇。
引用它的顶会 Paper12
- When XR and AI Meet - A Scoping Review on Extended Reality and Artificial IntelligenceTeresa Hirzle, Florian Müller, Fiona Draxler, Martin Schmitz 等CHI 2023 · 被引用 90 次
- ScalAR: Authoring Semantically Adaptive Augmented Reality Experiences in Virtual RealityXun Qian, Fengming He, Xiyun Hu, Tianyi Wang 等CHI 2022 · 被引用 69 次
- Voxel-based 3D Detection and Reconstruction of Multiple Objects from a Single ImageFeng Liu, Xiaoming LiuNeurIPS 2021 · 被引用 43 次
- Training an Open-Vocabulary Monocular 3D Detection Model without 3D DataRui Huang, Henry Zheng, Yan Wang, Zhuofan Xia 等NeurIPS 2024 · 被引用 26 次
- MineXR: Mining Personalized Extended Reality InterfacesHyunsung Cho, Yukang Yan, Kashyap Todi, Mark Parent 等CHI 2024 · 被引用 22 次
相关 Paper
- Interpolation-Aware Padding for 3D Sparse Convolutional Neural NetworksYu-Qi Yang, Peng-Shuai Wang, Yang LiuICCV 2021 · 被引用 4 次
- Not All Neighbors Matter: Point Distribution-Aware Pruning for 3D Point CloudYejin Lee, Donghyun Lee, JungUk Hong, Jae W. Lee 等AAAI 2023 · 被引用 7 次
- Fusion-Aware Point Convolution for Online Semantic 3D Scene SegmentationJiazhao Zhang, Chenyang Zhu, Lintao Zheng, Kai XuCVPR 2020
- Interpolated Convolutional Networks for 3D Point Cloud UnderstandingJiageng Mao, Xiaogang Wang, Hongsheng LiICCV 2019 · 被引用 241 次
- INS-Conv: Incremental Sparse Convolution for Online 3D SegmentationLeyao Liu, Tian Zheng, Yun-Jou Lin, Kai Ni 等CVPR 2022 · 被引用 19 次
