LSK3DNet: Towards Effective and Efficient 3D Perception with Large Sparse Kernels
Tuo Feng, Wenguan Wang, Fan Ma, Yi Yang
Abstract
Autonomous systems need to process large-scale, sparse, and irregular point clouds with limited compute resources. Consequently, it is essential to develop LiDAR perception methods that are both efficient and effective. Although naively enlarging 3D kernel size can enhance performance, it will also lead to a cubically-increasing overhead. Therefore, it is crucial to develop streamlined 3D large kernel designs that eliminate redundant weights and work effectively with larger kernels. In this paper, we propose an efficient and effective Large Sparse Kernel 3D Neural Network (LSK3DNet) that leverages dynamic pruning to amplify the 3D kernel size.
Our method comprises two core components: Spatial-wise Dynamic Sparsity (SDS) and Channel-wise Weight Selection (CWS). SDS dynamically prunes and regrows volumetric weights from the beginning to learn a large sparse 3D kernel. It not only boosts performance but also significantly reduces model size and computational cost. Moreover, CWS selects the most important channels for 3D convolution during training and subsequently prunes the redundant channels to accelerate inference for 3D vision tasks. We demonstrate the effectiveness of LSK3DNet on three benchmark datasets and five tracks compared with classical models and large kernel designs. Notably, LSK3DNet achieves the state-of-the-art performance on SemanticKITTI (i.e., 75.6% on single-scan and 63.4% on multi-scan), with roughly 40% model size reduction and 60% computing operations reduction compared to the naive large 3D kernel model.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 4e3b6153-fd26-4043-a720-34a123b0fce8Cited by top-tier papers5
- AdaSFormer: Adaptive Serialized Transformers for Monocular Semantic Scene Completion from Indoor EnvironmentsXuzhi Wang, Xinran Wu, Song Wang, Lingdong Kong et al.CVPR 2026 · 3 citations
- UniMapping: Unified SLAM Framework for Map-Centric Embodied PerceptionXiaze Zhang, Ziheng Ding, Yuejie Zhang, lifeng chen et al.ICML 2026
- Learning to Identify Out-of-Distribution Objects for 3D LiDAR Anomaly SegmentationSimone Mosco, Daniel Fusaro, Alberto PrettoCVPR 2026
- MFINet: Multi-view Fusion and 2D-3D Interaction Enhancement for Real-Time LiDAR Semantic SegmentationNan Ma, Zhijie Liu, Yiheng HanAAAI 2026
- Hierarchical Direction Perception via Atomic Dot-Product Operators for Rotation-Invariant Point Clouds LearningChenyu Hu, Xiaotong Li, Hao Zhu, Biao HouAAAI 2026
Builds on48
- Swin Transformer: Hierarchical Vision Transformer using Shifted WindowsZe Liu, Yutong Lin, Yue Cao, Han Hu et al.ICCV 2021 · 31,683 citations
- KPConv: Flexible and Deformable Convolution for Point CloudsHugues Thomas, Charles R. Qi, Jean-Emmanuel Deschaud, Beatriz Marcotegui et al.ICCV 2019 · 3,193 citations
- SemanticKITTI: A Dataset for Semantic Scene Understanding of LiDAR SequencesJens Behley, Martin Garbade, Andres Milioto, Jan Quenzel et al.ICCV 2019 · 2,345 citations
- Deep Hough Voting for 3D Object Detection in Point CloudsCharles R. Qi, Or Litany, Kaiming He, Leonidas J. GuibasICCV 2019 · 1,467 citations
- Voxel R-CNN: Towards High Performance Voxel-based 3D Object DetectionJiajun Deng, Shaoshuai Shi, Peiwei Li, Wengang Zhou et al.AAAI 2021 · 1,128 citations
Related papers
- LinK: Linear Kernel for LiDAR-based 3D PerceptionTao Lu, Xiang Ding, Haisong Liu, Gangshan Wu et al.CVPR 2023
- LargeKernel3D: Scaling up Kernels in 3D Sparse CNNsYukang Chen, Jianhui Liu, Xiangyu Zhang, Xiaojuan Qi et al.CVPR 2023
- Not All Neighbors Matter: Point Distribution-Aware Pruning for 3D Point CloudYejin Lee, Donghyun Lee, JungUk Hong, Jae W. Lee et al.AAAI 2023 · 7 citations
- Cylindrical and Asymmetrical 3D Convolution Networks for LiDAR SegmentationXinge Zhu, Hui Zhou, Tai Wang, Fangzhou Hong et al.CVPR 2021
- RandLA-Net: Efficient Semantic Segmentation of Large-Scale Point CloudsQingyong Hu, Bo Yang, Linhai Xie, Stefano Rosa et al.CVPR 2020
