Kecor: Kernel Coding Rate Maximization for Active 3D Object Detection
Yadan Luo, Zhuoxiao Chen, Zhen Fang, Zheng Zhang, Mahsa Baktashmotlagh, Zi Huang
摘要
Achieving a reliable LiDAR-based object detector in autonomous driving is paramount, but its success hinges on obtaining large amounts of precise 3D annotations. Active learning (AL) seeks to mitigate the annotation burden through algorithms that use fewer labels and can attain performance comparable to fully supervised learning. Although AL has shown promise, current approaches prioritize the selection of unlabeled point clouds with high uncertainty and/or diversity, leading to the selection of more instances for labeling and reduced computational efficiency. In this paper, we resort to a novel kernel coding rate maximization (Kecor) strategy which aims to identify the most informative point clouds to acquire labels through the lens of information theory. Greedy search is applied to seek desired point clouds that can maximize the minimal number of bits required to encode the latent features. To determine the uniqueness and informativeness of the selected samples from the model perspective, we construct a proxy network of the 3D detector head and compute the outer product of Jacobians from all proxy layers to form the empirical neural tangent kernel (NTK) matrix. To accommodate both one-stage (i.e., Second) and two-stage detectors (i.e., Pv-rcnn), we further incorporate the classification entropy maximization and well trade-off between detection performance and the total number of bounding boxes selected for annotation. Extensive experiments conducted on two 3D benchmarks and a 2D detection dataset evidence the superiority and versatility of the proposed approach. Our results show that approximately 44% box-level annotation costs and 26% computational time are reduced compared to the state-of-the-art AL method, without compromising detection performance. Source code: https://github.com/Luoyadan/KECOR-active-3Ddet.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了最后一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper7
- DPO: Dual-Perturbation Optimization for Test-time Adaptation in 3D Object DetectionZhuoxiao Chen, Zixin Wang, Yadan Luo, Sen Wang 等ACM MM 2024 · 被引用 3 次
- STONE: A Submodular Optimization Framework for Active 3D Object DetectionRuiyu Mao, Sarthak Kumar Maharana, Rishabh K. Iyer, Yunhui GuoNeurIPS 2024 · 被引用 3 次
- Advancing Prompt Learning through an External LayerFangming Cui, Xun Yang, Chao Wu, Liang Xiao 等ACM MM 2024 · 被引用 3 次
- CodeMerge: Codebook-Guided Model Merging for Robust Test-Time Adaptation in Autonomous DrivingHuitong Yang, Zhuoxiao Chen, Fengyi Zhang, Zi Huang 等NeurIPS 2025 · 被引用 2 次
- Uncertainty Meets Diversity: A Comprehensive Active Learning Framework for Indoor 3D Object DetectionJiangyi Wang, Na ZhaoCVPR 2025
它引用的顶会 Paper26
- Voxel R-CNN: Towards High Performance Voxel-based 3D Object DetectionJiajun Deng, Shaoshuai Shi, Peiwei Li, Wengang Zhou 等AAAI 2021 · 被引用 1,128 次
- Deep Batch Active Learning by Diverse, Uncertain Gradient Lower BoundsJordan T. Ash, Chicheng Zhang, Akshay Krishnamurthy, John Langford 等ICLR 2020 · 被引用 974 次
- STD: Sparse-to-Dense 3D Object Detector for Point CloudZetong Yang, Yanan Sun, Shu Liu, Xiaoyong Shen 等ICCV 2019 · 被引用 840 次
- Learning Diverse and Discriminative Representations via the Principle of Maximal Coding Rate ReductionYaodong Yu, Kwan Ho Ryan Chan, Chong You, Chaobing Song 等NeurIPS 2020 · 被引用 265 次
- Batch Active Learning at ScaleGui Citovsky, Giulia DeSalvo, Claudio Gentile, Lazaros Karydas 等NeurIPS 2021 · 被引用 220 次
相关 Paper
- Exploring Active 3D Object Detection from a Generalization PerspectiveYadan Luo, Zhuoxiao Chen, Zijian Wang, Xin Yu 等ICLR 2023 · 被引用 3 次
- ReDAL: Region-based and Diversity-aware Active Learning for Point Cloud Semantic SegmentationTsung-Han Wu, Yueh-Cheng Liu, Yu-Kai Huang, Hsin-Ying Lee 等ICCV 2021 · 被引用 92 次
- Plug and Play Active Learning for Object DetectionChenhongyi Yang, Lichao Huang, Elliot J. CrowleyCVPR 2024 · 被引用 29 次
- You Never Get a Second Chance To Make a Good First Impression: Seeding Active Learning for 3D Semantic SegmentationNermin Samet, Oriane Siméoni, Gilles Puy, Georgy Ponimatkin 等ICCV 2023 · 被引用 9 次
- VoxelKP: A Voxel-Based Network Architecture for Human Keypoint Estimation in LiDAR DataJian Shi, Peter WonkaICCV 2025
