AIDE: An Automatic Data Engine for Object Detection in Autonomous Driving
Mingfu Liang, Jong-Chyi Su, Samuel Schulter, Sparsh Garg, Shiyu Zhao, Ying Wu, Manmohan Chandraker
摘要
Autonomous vehicle (AV) systems rely on robust perception models as a cornerstone of safety assurance. However, objects encountered on the road exhibit a long-tailed distribution, with rare or unseen categories posing challenges to a deployed perception model. This necessitates an expensive process of continuously curating and annotating data with significant human effort. We propose to leverage recent advances in vision-language and large language models to design an Automatic Data Engine (AIDE) that automatically identifies issues, efficiently curates data, improves the model through auto-labeling, and verifies the model through generation of diverse scenarios. This process operates iteratively, allowing for continuous self-improvement of the model. We further establish a benchmark for open-world detection on AV datasets to comprehensively evaluate various learning paradigms, demonstrating our method's superior performance at a reduced cost.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper10
- ROADWork: A Dataset and Benchmark for Learning to Recognize, Observe, Analyze and Drive Through Work ZonesAnurag Ghosh, Shen Zheng, Robert Tamburo, Khiem Vuong 等ICCV 2025 · 被引用 4 次
- No Labels, No Problem: Training Visual Reasoners with Multimodal VerifiersDamiano Marsili, Georgia GkioxariICLR 2026 · 被引用 3 次
- AiDE-Q: Synthetic Labeled Datasets Can Enhance Learning Models for Quantum Property EstimationXinbiao Wang, Yuxuan Du, Zihan Lou, Yang Qian 等NeurIPS 2025 · 被引用 1 次
- Towards Scalable Spatial Intelligence Via 2D-To-3D Data LiftingXingyu Miao, Haoran Duan, Quanhao Qian, Jiuniu Wang 等ICCV 2025 · 被引用 1 次
- Can't Slow Me Down: Learning Robust and Hardware-Adaptive Object Detectors against Latency Attacks for Edge DevicesTianyi Wang, Zichen Wang, Cong Wang, Yuanchao Shu 等CVPR 2025
它引用的顶会 Paper32
- Learning Transferable Visual Models From Natural Language SupervisionAlec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh 等ICML 2021 · 被引用 47,906 次
- Segment AnythingAlexander Kirillov, Eric Mintun, Nikhila Ravi, Hanzi Mao 等ICCV 2023 · 被引用 13,211 次
- BLIP-2: Bootstrapping Language-Image Pre-training with Frozen Image Encoders and Large Language ModelsJunnan Li, Dongxu Li, Silvio Savarese, Steven C. H. HoiICML 2023 · 被引用 7,873 次
- Open-vocabulary Object Detection via Vision and Language Knowledge DistillationXiuye Gu, Tsung-Yi Lin, Weicheng Kuo, Yin CuiICLR 2022 · 被引用 1,274 次
- Unbiased Teacher for Semi-Supervised Object DetectionYen-Cheng Liu, Chih-Yao Ma, Zijian He, Chia-Wen Kuo 等ICLR 2021 · 被引用 603 次
相关 Paper
- Lifting Unlabeled Internet-level Data for 3D Scene UnderstandingYixin Chen, Yaowei Zhang, Huangyue Yu, Junchao He 等CVPR 2026 · 被引用 1 次
- VILTA: A VLM-in-the-Loop Adversary for Enhancing Driving Policy RobustnessQimao Chen, Fang Li, Shaoqing Xu, Zhiyi Lai 等AAAI 2026 · 被引用 2 次
- Percept-WAM: Perception-Enhanced World-Awareness-Action Model for Robust End-to-End Autonomous DrivingJianhua Han, Meng Tian, Jiangtong Zhu, Fan He 等CVPR 2026 · 被引用 10 次
- The Blind Spot of Adaptation: Quantifying and Mitigating Forgetting in Fine-tuned Driving ModelsRunhao Mao, Hanshi Wang, Yixiang Yang, Qianli Ma 等CVPR 2026 · 被引用 1 次
- OD-RASE: Ontology-Driven Risk Assessment and Safety Enhancement for Autonomous DrivingKota Shimomura, Masaki Nambata, Atsuya Ishikawa, Ryota Mimura 等ICCV 2025 · 被引用 1 次
