MEDIRL: Predicting the Visual Attention of Drivers via Maximum Entropy Deep Inverse Reinforcement Learning
Sonia Baee, Erfan Pakdamanian, Inki Kim, Lu Feng, Vicente Ordonez, Laura E. Barnes
摘要
Inspired by human visual attention, we propose a novel inverse reinforcement learning formulation using Maximum Entropy Deep Inverse Reinforcement Learning (MEDIRL) for predicting the visual attention of drivers in accident-prone situations. MEDIRL predicts fixation locations that lead to maximal rewards by learning a task-sensitive reward function from eye fixation patterns recorded from attentive drivers. Additionally, we introduce EyeCar, a new driver attention dataset in accident-prone situations. We conduct comprehensive experiments to evaluate our proposed model on three common benchmarks: (DR(eye)VE, BDD-A, DADA-2000), and our EyeCar dataset. Results indicate that MEDIRL outperforms existing models for predicting attention and achieves state-of-the-art performance. We present extensive ablation studies to provide more insights into different features of our proposed model.1
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了最后一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper10
- Unsupervised Self-Driving Attention Prediction via Uncertainty Mining and Knowledge EmbeddingPengfei Zhu, Mengshi Qi, Xia Li, Weijian Li 等ICCV 2023 · 被引用 23 次
- FBLNet: FeedBack Loop Network for Driver Attention PredictionYilong Chen, Zhixiong Nan, Tao XiangICCV 2023 · 被引用 21 次
- SalM²: An Extremely Lightweight Saliency Mamba Model for Real-Time Cognitive Awareness of Driver AttentionChunyu Zhao, Wentao Mu, Xian Zhou, Wenbo Liu 等AAAI 2025 · 被引用 16 次
- Where, What, Why: Towards Explainable Driver Attention PredictionYuchen Zhou, Jiayu Tang, Xiaoyan Xiao, Yueyao Lin 等ICCV 2025 · 被引用 8 次
- Cross-Modality Graph-based Language and Sensor Data Co-Learning of Human-Mobility InteractionMahan Tabatabaie, Suining He, Kang G. ShinUbiComp 2023 · 被引用 6 次
它引用的顶会 Paper11
- Digging Into Self-Supervised Monocular Depth EstimationClément Godard, Oisin Mac Aodha, Michael Firman, Gabriel J. BrostowICCV 2019 · 被引用 2,416 次
- Exploring the Limitations of Behavior Cloning for Autonomous DrivingFelipe Codevilla, Eder Santana, Antonio M. López, Adrien GaidonICCV 2019 · 被引用 666 次
- Video Instance SegmentationLinjie Yang, Yuchen Fan, Ning XuICCV 2019 · 被引用 615 次
- TASED-Net: Temporally-Aggregating Spatial Encoder-Decoder Network for Video Saliency DetectionKyle Min, Jason J. CorsoICCV 2019 · 被引用 189 次
- DGaze: CNN-Based Gaze Prediction in Dynamic ScenesZhiming Hu, Sheng Li, Congyi Zhang, Kangrui Yi 等IEEE VR 2020 · 被引用 105 次
相关 Paper
- DRIVE: Deep Reinforced Accident Anticipation with Visual ExplanationWentao Bao, Qi Yu, Yu KongICCV 2021 · 被引用 64 次
- Predicting Goal-Directed Human Attention Using Inverse Reinforcement LearningZhibo Yang, Lihan Huang, Yupei Chen, Zijun Wei 等CVPR 2020
- GoIRL: Graph-Oriented Inverse Reinforcement Learning for Multimodal Trajectory PredictionMuleilan Pei, Shaoshuai Shi, Lu Zhang, Peiliang Li 等ICML 2025
- From Gaze to Movement: Predicting Visual Attention for Autonomous Driving Human-Machine Interaction based on Programmatic Imitation LearningYexin Huang, Yongbin Lin, Lishengsa Yue, Zhihong Yao 等ICCV 2025 · 被引用 2 次
- PRE-MAP: Personalized Reinforced Eye-tracking Multimodal LLM for High-Resolution Multi-Attribute Point PredictionHanbing Wu, Ping Jiang, Anyang Su, Chenxu Zhao 等ACM MM 2025
