xGAIL: Explainable Generative Adversarial Imitation Learning for Explainable Human Decision Analysis
Menghai Pan, Weixiao Huang, Yanhua Li, Xun Zhou, Jun Luo
摘要
To make daily decisions, human agents devise their own "strategies" governing their mobility dynamics (e.g., taxi drivers have preferred working regions and times, and urban commuters have preferred routes and transit modes). Recent research such as generative adversarial imitation learning (GAIL) demonstrates successes in learning human decision-making strategies from their behavior data using deep neural networks (DNNs), which can accurately mimic how humans behave in various scenarios, e.g., playing video games, etc. However, such DNN-based models are "black box" models in nature, making it hard to explain what knowledge the models have learned from human, and how the models make such decisions, which was not addressed in the literature of imitation learning. This paper addresses this research gap by proposing xGAIL, the first explainable generative adversarial imitation learning framework. The proposed xGAIL framework consists of two novel components, including Spatial Activation Maximization (SpatialAM) and Spatial Randomized Input Sampling Explanation (SpatialRISE), to extract both global and local knowledge from a well-trained GAIL model that explains how a human agent makes decisions. Especially, we take taxi drivers' passenger-seeking strategy as an example to validate the effectiveness of the proposed xGAIL framework. Our analysis on a large-scale real-world taxi trajectory data shows promising results from two aspects: i) global explainable knowledge of what nearby traffic condition impels a taxi driver to choose a particular direction to find the next passenger, and ii) local explainable knowledge of what key (sometimes hidden) factors a taxi driver considers when making a particular decision.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper2
- ProtoX: Explaining a Reinforcement Learning Agent via PrototypingRonilo J. Ragodos, Tong Wang, Qihang Lin, Xun ZhouNeurIPS 2022 · 被引用 13 次
- Urban-Focused Multi-Task Offline Reinforcement Learning with Contrastive Data SharingXinbo Zhao, Yingxue Zhang, Xin Zhang, Yu Yang 等KDD 2024 · 被引用 2 次
相关 Paper
- How Do We Move: Modeling Human Movement with System DynamicsHua Wei, Dongkuan Xu, Junjie Liang, Zhenhui LiAAAI 2021 · 被引用 16 次
- DecompGAIL: Learning Realistic Traffic Behaviors with Decomposed Multi-Agent Generative Adversarial Imitation LearningKe Guo, Haochen Liu, Xiaojun Wu, Chen LvICLR 2026 · 被引用 8 次
- Spotting Deep Neural Network Vulnerabilities in Mobile Traffic Forecasting with an Explainable AI LensSerly Moghadas, Claudio Fiandrino, Alan Collet, Giulia Attanasio 等INFOCOM 2023 · 被引用 9 次
- Reinforced Imitative Graph Representation Learning for Mobile User Profiling: An Adversarial Training PerspectiveDongjie Wang, Pengyang Wang, Kunpeng Liu, Yuanchun Zhou 等AAAI 2021 · 被引用 31 次
- Learning to Simulate Daily Activities via Modeling Dynamic Human NeedsYuan Yuan, Huandong Wang, Jingtao Ding, Depeng Jin 等WWW 2023 · 被引用 43 次
