Learning When Agents Can Talk to Drivers Using the INAGT Dataset and Multisensor Fusion
Tong Wu, Nikolas Martelaro, Simon Stent, Jorge Ortiz, Wendy Ju
摘要
This paper examines sensor fusion techniques for modeling opportunities for proactive speech-based in-car interfaces. We leverage the Is Now a Good Time (INAGT) dataset, which consists of automotive, physiological, and visual data collected from drivers who self-annotated responses to the question "Is now a good time?," indicating the opportunity to receive non-driving information during a 50-minute drive. We augment this original driver-annotated data with third-party annotations of perceived safety, in order to explore potential driver overconfidence. We show that fusing automotive, physiological, and visual data allows us to predict driver labels of availability, achieving an 0.874 F1-score by extracting statistically relevant features and training with our proposed deep neural network, PazNet. Using the same data and network, we achieve an 0.891 F1-score for predicting third-party labeled safe moments. We train these models to avoid false positives---determinations that it is a good time to interrupt when it is not---since false positives may cause driver distraction or service deactivation by the driver. Our analyses show that conservative models still leave many moments for interaction and show that most inopportune moments are short. This work lays a foundation for using sensor fusion models to predict when proactive speech systems should engage with drivers.
问问这篇 Paper
问问你的智能体。
Lune 读过与它相关的顶会 Paper,每个回答都会注明依据哪几篇。
引用它的顶会 Paper3
- AIDE: A Vision-Driven Multi-View, Multi-Modal, Multi-Tasking Dataset for Assistive Driving PerceptionDingkang Yang, Shuai Huang, Zhi Xu, Zhenpeng Li 等ICCV 2023 · 被引用 72 次
- RouteLLM: A Large Language Model with Native Route Context Understanding to Enable Context-Aware ReasoningPhilipp Hallgarten, Verena Jasmin Hallitschke, Enkelejda Kasneci, Michael Beigl 等UbiComp 2025 · 被引用 5 次
- ProVoice: Designing Proactive Functionality for In-Vehicle Conversational Assistants using Multi-Objective Bayesian Optimization to Enhance Driver ExperienceJosh Susak, Yifu Liu, Pascal Jansen, Mark ColleyCHI 2026 · 被引用 2 次
相关 Paper
- DeepTake: Prediction of Driver Takeover Behavior using Multimodal DataErfan Pakdamanian, Shili Sheng, Sonia Baee, Seongkook Heo 等CHI 2021 · 被引用 86 次
- TimelyTale: A Multimodal Dataset Approach to Assessing Passengers' Explanation Demands in Highly Automated VehiclesGwangbin Kim, Seokhyun Hwang, Minwoo Seong, Dohyeon Yeo 等UbiComp 2024 · 被引用 13 次
- Interruptibility for In-vehicle Multitasking: Influence of Voice Task Demands and Adaptive BehaviorsAuk Kim, Jung-Mi Park, Uichin LeeUbiComp 2020 · 被引用 30 次
- Small Talk, Big Impact? LLM-based Conversational Agents to Mitigate Passive Fatigue in Conditional Automated DrivingLewis Cockram, Yueteng Yu, Jorge Pardo, Xiaomeng Li 等CHI 2026 · 被引用 3 次
- From Awareness to Intent: Mitigating Silent Driving System Failures through Prospective Situation Awareness Enhancing InterfacesJiyao Wang, Song Yan, Xiao Yang, Qihang He 等CHI 2026 · 被引用 4 次
