Curricular Object Manipulation in LiDAR-based Object Detection
Ziyue Zhu, Qiang Meng, Xiao Wang, Ke Wang, Liujiang Yan, Jian Yang
摘要
This paper explores the potential of curriculum learning in LiDAR-based 3D object detection by proposing a curricular object manipulation (COM) framework. The framework embeds the curricular training strategy into both the loss design and the augmentation process. For the loss design, we propose the COMLoss to dynamically predict object-level difficulties and emphasize objects of different difficulties based on training stages. On top of the widely-used augmentation technique called GT-Aug in Li-DAR detection tasks, we propose a novel COMAug strategy which first clusters objects in ground-truth database based on well-designed heuristics. Group-level difficulties rather than individual ones are then predicted and updated during training for stable results. Model performance and generalization capabilities can be improved by sampling and augmenting progressively more difficult objects into the training samples. Extensive experiments and ablation studies reveal the superior and generality of the proposed framework. The code is available at https://github.com/ZZY816/COM .
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper6
- IS-Fusion: Instance-Scene Collaborative Fusion for Multimodal 3D Object DetectionJunbo Yin, Jianbing Shen, Runnan Chen, Wei Li 等CVPR 2024 · 被引用 73 次
- OPUS: Occupancy Prediction Using a Sparse SetJiabao Wang, Zhaojiang Liu, Qiang Meng, Liujiang Yan 等NeurIPS 2024 · 被引用 67 次
- Reconstruction-Guided Slot Curriculum: Addressing Object Over-Fragmentation in Video Object-Centric LearningWonJun Moon, Hyun Seok Seong, Jae-Pil HeoCVPR 2026 · 被引用 3 次
- Completion as Enhancement: A Degradation-Aware Selective Image Guided Network for Depth CompletionZhiqiang Yan, Zhengxue Wang, Kun Wang, Jun Li 等CVPR 2025
- Learning 3D Perception from Others' PredictionsJinsu Yoo, Zhenyang Feng, Tai-Yu Pan, Yihong Sun 等ICLR 2025
它引用的顶会 Paper30
- An Image is Worth 16x16 Words: Transformers for Image Recognition at ScaleAlexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn 等ICLR 2021 · 被引用 21,477 次
- CenterNet: Keypoint Triplets for Object DetectionKaiwen Duan, Song Bai, Lingxi Xie, Honggang Qi 等ICCV 2019 · 被引用 3,348 次
- Generalized Focal Loss: Learning Qualified and Distributed Bounding Boxes for Dense Object DetectionXiang Li, Wenhai Wang, Lijun Wu, Shuo Chen 等NeurIPS 2020 · 被引用 2,118 次
- Deep Hough Voting for 3D Object Detection in Point CloudsCharles R. Qi, Or Litany, Kaiming He, Leonidas J. GuibasICCV 2019 · 被引用 1,467 次
- Voxel R-CNN: Towards High Performance Voxel-based 3D Object DetectionJiajun Deng, Shaoshuai Shi, Peiwei Li, Wengang Zhou 等AAAI 2021 · 被引用 1,128 次
相关 Paper
- LiDAR-Aug: A General Rendering-Based Augmentation Framework for 3D Object DetectionJin Fang, Xinxin Zuo, Dingfu Zhou, Shengze Jin 等CVPR 2021
- Just Add $100 More: Augmenting Pseudo-LiDAR Point Cloud for Resolving Class-imbalance ProblemMincheol Chang, Siyeong Lee, Jinkyu Kim, Namil KimNeurIPS 2024 · 被引用 4 次
- Exploring Geometry-aware Contrast and Clustering Harmonization for Self-supervised 3D Object DetectionHanxue Liang, Chenhan Jiang, Dapeng Feng, Xin Chen 等ICCV 2021 · 被引用 85 次
- Pre-training LiDAR-based 3D Object Detectors through ColorizationTai-Yu Pan, Chenyang Ma, Tianle Chen, Cheng Perng Phoo 等ICLR 2024 · 被引用 5 次
- Exploring Active 3D Object Detection from a Generalization PerspectiveYadan Luo, Zhuoxiao Chen, Zijian Wang, Xin Yu 等ICLR 2023 · 被引用 3 次
