Sniffing Threatening Open-World Objects in Autonomous Driving by Open-Vocabulary Models
Yulin He, Siqi Wang, Wei Chen, Tianci Xun, Yusong Tan
摘要
Autonomous driving (AD) is a typical application that requires effectively exploiting multimedia information. For AD, it is critical to ensure safety by detecting unknown objects in an open world, driving the demand for open world object detection (OWOD). However, existing OWOD methods treat generic objects beyond known classes in the train set as unknown objects and prioritize recall in evaluation. This encourages excessive false positives and endangers safety of AD. To address this issue, we restrict the definition of unknown objects to threatening objects in AD, and introduce a new evaluation protocol, which is built upon a new metric named U-ARecall, to alleviate biased evaluation caused by neglecting false positives. Under the new evaluation protocol, we re-evaluate existing OWOD methods and discover that they typically perform poorly in AD. Then, we propose a novel OWOD paradigm for AD based on fine-tuning foundational open-vocabulary models (OVMs), as they can exploit rich linguistic and visual prior knowledge for OWOD. Following this new paradigm, we propose a brand-new OWOD solution, which effectively addresses two core challenges of fine-tuning OVMs via two novel techniques: 1) the maintenance of open-world generic knowledge by a dual-branch architecture; 2) the acquisition of scenario-specific knowledge by the visual-oriented contrastive learning scheme. Besides, a dual-branch prediction fusion module is proposed to avoid post-processing and hand-crafted heuristics. Extensive experiments show that our proposed method not only surpasses classic OWOD methods in unknown object detection by a large margin (∼× U-ARecall), but also notably outperforms OVMs without fine-tuning in known object detection (∼ 20% K-mAP). Our codes are available at https://github.com/harrylin-hyl/AD-OWOD.
问问这篇 Paper
问问你的智能体。
Lune 读过与它相关的顶会 Paper,每个回答都会注明依据哪几篇。
引用它的顶会 Paper2
- NavigScene: Bridging Local Perception and Global Navigation for Beyond-Visual-Range Autonomous DrivingQucheng Peng, Chen Bai, Guoxiang Zhang, Bo Xu 等ACM MM 2025 · 被引用 6 次
- Attention to Threat-Relevant Objects: Reasoning Detection in Autonomous Driving via Multimodal Large Language ModelsYulin He, Wei Chen, Xinbiao Gan, Siqi Wang 等AAAI 2026
相关 Paper
- OW-Adapter: Human-Assisted Open-World Object Detection with a Few ExamplesSuphanut Jamonnak, Jiajing Guo, Wenbin He, Liang Gou 等IEEE VIS 2023 · 被引用 6 次
- Rethinking Open-World Object Detection in Autonomous Driving ScenariosZeyu Ma, Yang Yang, Guoqing Wang, Xing Xu 等ACM MM 2022 · 被引用 39 次
- Hyp-OW: Exploiting Hierarchical Structure Learning with Hyperbolic Distance Enhances Open World Object DetectionThang Doan, Xin Li, Sima Behpour, Wenbin He 等AAAI 2024 · 被引用 19 次
- OW-DETR: Open-world Detection TransformerAkshita Gupta, Sanath Narayan, K. J. Joseph, Salman Khan 等CVPR 2022 · 被引用 209 次
- Open-Scenario Domain Adaptive Object Detection in Autonomous DrivingZeyu Ma, Ziqiang Zheng, Jiwei Wei, Xiaoyong Wei 等ACM MM 2023 · 被引用 2 次
