Secrets Everywhere: Auditing Memorization in Mobility Prediction Models
Anne Josiane Kouam, Hristo Boyadzhiev, Konrad Rieck
摘要
Human mobility prediction models, which forecast the next location in a user's trajectory, are increasingly deployed in urban analytics, navigation, and personalized services. Yet, little is known about their potential to memorize and expose sensitive user trajectories from training data. While memorization has been extensively studied in language models, mobility prediction poses unique challenges: training sequences encode human behavior at various spatial and temporal scales, creating privacy risks at different granularities. In this paper, we conduct the first systematic audit of memorization in mobility prediction models. While prior work has shown that privacy leaks can arise from such models, we systematically assess and quantify memorization risks at scale. We identify key challenges, including the lack of a randomness space, the multi-scale structure of trajectories, and user-specific behavioral diversity. To address these challenges, we introduce a framework to quantify mobility memorization at different levels of granularity: individual locations, anchor pairs, and subtrajectory segments. We also develop user-grounded reference sets to assess how likely a model is to prefer training data over realistic alternatives. Our evaluation across multiple models and datasets reveals pervasive memorization patterns that correlate with user regularity and increase the risk of data extraction at inference time. Our findings call for mandatory privacy auditing in mobility prediction models.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper7
- Extracting Training Data from Large Language ModelsNicholas Carlini, Florian Tramèr, Eric Wallace, Matthew Jagielski 等USENIX Security 2021 · 被引用 2,866 次
- The Secret Sharer: Evaluating and Testing Unintended Memorization in Neural NetworksNicholas Carlini, Chang Liu, Úlfar Erlingsson, Jernej Kos 等USENIX Security 2019 · 被引用 1,386 次
- Deduplicating Training Data Makes Language Models BetterKatherine Lee, Daphne Ippolito, Andrew Nystrom, Chiyuan Zhang 等ACL 2022 · 被引用 844 次
- Where to Go Next: Modeling Long- and Short-Term User Preferences for Point-of-Interest RecommendationKe Sun, Tieyun Qian, Tong Chen, Yile Liang 等AAAI 2020 · 被引用 412 次
- Knock Knock, Who's There? Membership Inference on Aggregate Location DataApostolos Pyrgelis, Carmela Troncoso, Emiliano De CristofaroNDSS 2018 · 被引用 293 次
相关 Paper
- Where Have You Been? A Study of Privacy Risk for Point-of-Interest RecommendationKunlin Cai, Jinghuai Zhang, Zhiqing Hong, William Shand 等KDD 2024 · 被引用 5 次
- TransRisk: Mobility Privacy Risk Prediction based on Transferred KnowledgeXiaoyang Xie, Zhiqing Hong, Zhou Qin, Zhihan Fang 等UbiComp 2022 · 被引用 1 次
- COLA: Cross-city Mobility Transformer for Human Trajectory SimulationYu Wang, Tongya Zheng, Yuxuan Liang, Shunyu Liu 等WWW 2024 · 被引用 37 次
- Going Where, by Whom, and at What Time: Next Location Prediction Considering User Preference and Temporal RegularityTianao Sun, Ke Fu, Weiming Huang, Kai Zhao 等KDD 2024 · 被引用 11 次
- AttnMove: History Enhanced Trajectory Recovery via Attentional NetworkTong Xia, Yunhan Qi, Jie Feng, Fengli Xu 等AAAI 2021 · 被引用 74 次
