Learning When Agents Can Talk to Drivers Using the INAGT Dataset and Multisensor Fusion
Tong Wu, Nikolas Martelaro, Simon Stent, Jorge Ortiz, Wendy Ju
Abstract
This paper examines sensor fusion techniques for modeling opportunities for proactive speech-based in-car interfaces. We leverage the Is Now a Good Time (INAGT) dataset, which consists of automotive, physiological, and visual data collected from drivers who self-annotated responses to the question "Is now a good time?," indicating the opportunity to receive non-driving information during a 50-minute drive. We augment this original driver-annotated data with third-party annotations of perceived safety, in order to explore potential driver overconfidence. We show that fusing automotive, physiological, and visual data allows us to predict driver labels of availability, achieving an 0.874 F1-score by extracting statistically relevant features and training with our proposed deep neural network, PazNet. Using the same data and network, we achieve an 0.891 F1-score for predicting third-party labeled safe moments. We train these models to avoid false positives---determinations that it is a good time to interrupt when it is not---since false positives may cause driver distraction or service deactivation by the driver. Our analyses show that conservative models still leave many moments for interaction and show that most inopportune moments are short. This work lays a foundation for using sensor fusion models to predict when proactive speech systems should engage with drivers.
Ask about this paper
Ask your agent about it.
Lune has read the top-tier papers around this one, so every answer names the papers it rests on.
Cited by top-tier papers3
- AIDE: A Vision-Driven Multi-View, Multi-Modal, Multi-Tasking Dataset for Assistive Driving PerceptionDingkang Yang, Shuai Huang, Zhi Xu, Zhenpeng Li et al.ICCV 2023 · 72 citations
- RouteLLM: A Large Language Model with Native Route Context Understanding to Enable Context-Aware ReasoningPhilipp Hallgarten, Verena Jasmin Hallitschke, Enkelejda Kasneci, Michael Beigl et al.UbiComp 2025 · 5 citations
- ProVoice: Designing Proactive Functionality for In-Vehicle Conversational Assistants using Multi-Objective Bayesian Optimization to Enhance Driver ExperienceJosh Susak, Yifu Liu, Pascal Jansen, Mark ColleyCHI 2026 · 2 citations
Related papers
- DeepTake: Prediction of Driver Takeover Behavior using Multimodal DataErfan Pakdamanian, Shili Sheng, Sonia Baee, Seongkook Heo et al.CHI 2021 · 86 citations
- TimelyTale: A Multimodal Dataset Approach to Assessing Passengers' Explanation Demands in Highly Automated VehiclesGwangbin Kim, Seokhyun Hwang, Minwoo Seong, Dohyeon Yeo et al.UbiComp 2024 · 13 citations
- Interruptibility for In-vehicle Multitasking: Influence of Voice Task Demands and Adaptive BehaviorsAuk Kim, Jung-Mi Park, Uichin LeeUbiComp 2020 · 30 citations
- Small Talk, Big Impact? LLM-based Conversational Agents to Mitigate Passive Fatigue in Conditional Automated DrivingLewis Cockram, Yueteng Yu, Jorge Pardo, Xiaomeng Li et al.CHI 2026 · 3 citations
- From Awareness to Intent: Mitigating Silent Driving System Failures through Prospective Situation Awareness Enhancing InterfacesJiyao Wang, Song Yan, Xiao Yang, Qihang He et al.CHI 2026 · 4 citations
