Towards Efficient 3D Object Detection with Knowledge Distillation
Jihan Yang, Shaoshuai Shi, Runyu Ding, Zhe Wang, Xiaojuan Qi
摘要
Despite substantial progress in 3D object detection, advanced 3D detectors often suffer from heavy computation overheads. To this end, we explore the potential of knowledge distillation (KD) for developing efficient 3D object detectors, focusing on popular pillar- and voxel-based detectors.In the absence of well-developed teacher-student pairs, we first study how to obtain student models with good trade offs between accuracy and efficiency from the perspectives of model compression and input resolution reduction. Then, we build a benchmark to assess existing KD methods developed in the 2D domain for 3D object detection upon six well-constructed teacher-student pairs. Further, we propose an improved KD pipeline incorporating an enhanced logit KD method that performs KD on only a few pivotal positions determined by teacher classification response, and a teacher-guided student model initialization to facilitate transferring teacher model's feature extraction ability to students through weight inheritance. Finally, we conduct extensive experiments on the Waymo dataset. Our best performing model achieves LEVEL 2 mAPH, surpassing its teacher model and requiring only of teacher flops. Our most efficient model runs 51 FPS on an NVIDIA A100, which is faster than PointPillar with even higher accuracy. Code is available at https://github.com/CVMI-Lab/SparseKD.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper17
- Not All Voxels are Equal: Hardness-Aware Semantic Scene Completion with Self-DistillationSong Wang, Jiawei Yu, Wentong Li, Wenyu Liu 等CVPR 2024 · 被引用 22 次
- Radio2Text: Streaming Speech Recognition Using mmWave Radio SignalsRunning Zhao, Jiangtao Yu, Hang Zhao, Edith C. H. NgaiUbiComp 2023 · 被引用 22 次
- CRKD: Enhanced Camera-Radar Object Detection with Cross-Modality Knowledge DistillationLingjun Zhao, Jingyu Song, Katherine A. SkinnerCVPR 2024 · 被引用 21 次
- Autoencoders as Cross-Modal Teachers: Can Pretrained 2D Image Transformers Help 3D Representation Learning?Runpei Dong, Zekun Qi, Linfeng Zhang, Junbo Zhang 等ICLR 2023 · 被引用 21 次
- STXD: Structural and Temporal Cross-Modal Distillation for Multi-View 3D Object DetectionSujin Jang, Dae Ung Jo, Sung Ju Hwang, Dongwook Lee 等NeurIPS 2023 · 被引用 19 次
它引用的顶会 Paper22
- Deformable DETR: Deformable Transformers for End-to-End Object DetectionXizhou Zhu, Weijie Su, Lewei Lu, Bin Li 等ICLR 2021 · 被引用 7,353 次
- Once-for-All: Train One Network and Specialize it for Efficient DeploymentHan Cai, Chuang Gan, Tianzhe Wang, Zhekai Zhang 等ICLR 2020 · 被引用 1,522 次
- Voxel R-CNN: Towards High Performance Voxel-based 3D Object DetectionJiajun Deng, Shaoshuai Shi, Peiwei Li, Wengang Zhou 等AAAI 2021 · 被引用 1,128 次
- STD: Sparse-to-Dense 3D Object Detector for Point CloudZetong Yang, Yanan Sun, Shu Liu, Xiaoyong Shen 等ICCV 2019 · 被引用 840 次
- Learning Lightweight Lane Detection CNNs by Self Attention DistillationYuenan Hou, Zheng Ma, Chunxiao Liu, Chen Change LoyICCV 2019 · 被引用 666 次
相关 Paper
- PointDistiller: Structured Knowledge Distillation Towards Efficient and Compact 3D DetectionLinfeng Zhang, Runpei Dong, Hung-Shuo Tai, Kaisheng MaCVPR 2023
- Representation Disparity-aware Distillation for 3D Object DetectionYanjing Li, Sheng Xu, Mingbao Lin, Jihao Yin 等ICCV 2023 · 被引用 6 次
- itKD: Interchange Transfer-based Knowledge Distillation for 3D Object DetectionHyeon Cho, Junyong Choi, Geonwoo Baek, Wonjun HwangCVPR 2023
- Joint Homophily and Heterophily Relational Knowledge Distillation for Efficient and Compact 3D Object DetectionShidi Chen, Lili Wei, Liqian Liang, Congyan LangACM MM 2024 · 被引用 1 次
- CaKDP: Category-Aware Knowledge Distillation and Pruning Framework for Lightweight 3D Object DetectionHaonan Zhang, Longjun Liu, Yuqi Huang, Zhao Yang 等CVPR 2024 · 被引用 10 次
