On-Road Object Importance Estimation: A New Dataset and A Model with Multi-Fold Top-Down Guidance
Zhixiong Nan, Yilong Chen, Tianfei Zhou, Tao Xiang
摘要
This paper addresses the problem of on-road object importance estimation, which utilizes video sequences captured from the driver's perspective as the input. Although this problem is significant for safer and smarter driving systems, the exploration of this problem remains limited. On one hand, publicly-available large-scale datasets are scarce in the community. To address this dilemma, this paper contributes a new large-scale dataset named Traffic Object Importance (TOI). On the other hand, existing methods often only consider either bottom-up feature or single-fold guidance, leading to limitations in handling highly dynamic and diverse traffic scenarios. Different from existing methods, this paper proposes a model that integrates multi-fold top-down guidance with the bottom-up feature. Specifically, three kinds of top-down guidance factors (i.e., driver intention, semantic context, and traffic rule) are integrated into our model. These factors are important for object importance estimation, but none of the existing methods simultaneously consider them. To our knowledge, this paper proposes the first on-road object importance estimation model that fuses multi-fold top-down guidance factors with bottom-up feature. Extensive experiments demonstrate that our model outperforms state-of-the-art methods by large margins, achieving 23.1% Average Precision (AP) improvement compared with the recently proposed model (i.e., Goal).
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper8
- CLRNet: Cross Layer Refinement Network for Lane DetectionTu Zheng, Yifei Huang, Yang Liu, Wenjian Tang 等CVPR 2022 · 被引用 280 次
- Pyramid Grafting Network for One-Stage High Resolution Saliency DetectionChenxi Xie, Changqun Xia, Mingcan Ma, Zhirui Zhao 等CVPR 2022 · 被引用 112 次
- FBLNet: FeedBack Loop Network for Driver Attention PredictionYilong Chen, Zhixiong Nan, Tao XiangICCV 2023 · 被引用 21 次
- nuScenes: A Multimodal Dataset for Autonomous DrivingHolger Caesar, Varun Bankiti, Alex H. Lang, Sourabh Vora 等CVPR 2020
- Learning to Detect Important People in Unlabelled Images for Semi-Supervised Important People DetectionFa-Ting Hong, Wei-Hong Li, Wei-Shi ZhengCVPR 2020
相关 Paper
- DriverGaze360: OmniDirectional Driver Attention with Object-Level GuidanceShreedhar Govil, Didier Stricker, Jason R. RambachCVPR 2026
- LOKI: Long Term and Key Intentions for Trajectory PredictionHarshayu Girase, Haiming Gang, Srikanth Malla, Jiachen Li 等ICCV 2021 · 被引用 102 次
- TrafficMOT: A Challenging Dataset for Multi-Object Tracking in Complex Traffic ScenariosLihao Liu, Yanqi Cheng, Zhongying Deng, Shujun Wang 等ACM MM 2024 · 被引用 2 次
- TITAN: Future Forecast Using Action PriorsSrikanth Malla, Behzad Dariush, Chiho ChoiCVPR 2020
- CRASH: Crash Recognition and Anticipation System Harnessing with Context-Aware and Temporal Focus AttentionsHaicheng Liao, Haoyu Sun, Huanming Shen, Chengyue Wang 等ACM MM 2024 · 被引用 10 次
