Secrets Everywhere: Auditing Memorization in Mobility Prediction Models
Anne Josiane Kouam, Hristo Boyadzhiev, Konrad Rieck
Abstract
Human mobility prediction models, which forecast the next location in a user's trajectory, are increasingly deployed in urban analytics, navigation, and personalized services. Yet, little is known about their potential to memorize and expose sensitive user trajectories from training data. While memorization has been extensively studied in language models, mobility prediction poses unique challenges: training sequences encode human behavior at various spatial and temporal scales, creating privacy risks at different granularities. In this paper, we conduct the first systematic audit of memorization in mobility prediction models. While prior work has shown that privacy leaks can arise from such models, we systematically assess and quantify memorization risks at scale. We identify key challenges, including the lack of a randomness space, the multi-scale structure of trajectories, and user-specific behavioral diversity. To address these challenges, we introduce a framework to quantify mobility memorization at different levels of granularity: individual locations, anchor pairs, and subtrajectory segments. We also develop user-grounded reference sets to assess how likely a model is to prefer training data over realistic alternatives. Our evaluation across multiple models and datasets reveals pervasive memorization patterns that correlate with user regularity and increase the risk of data extraction at inference time. Our findings call for mandatory privacy auditing in mobility prediction models.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext d2df3029-c6fb-417e-8820-e8fb11649595Builds on7
- Extracting Training Data from Large Language ModelsNicholas Carlini, Florian Tramèr, Eric Wallace, Matthew Jagielski et al.USENIX Security 2021 · 2,866 citations
- The Secret Sharer: Evaluating and Testing Unintended Memorization in Neural NetworksNicholas Carlini, Chang Liu, Úlfar Erlingsson, Jernej Kos et al.USENIX Security 2019 · 1,386 citations
- Deduplicating Training Data Makes Language Models BetterKatherine Lee, Daphne Ippolito, Andrew Nystrom, Chiyuan Zhang et al.ACL 2022 · 844 citations
- Where to Go Next: Modeling Long- and Short-Term User Preferences for Point-of-Interest RecommendationKe Sun, Tieyun Qian, Tong Chen, Yile Liang et al.AAAI 2020 · 412 citations
- Knock Knock, Who's There? Membership Inference on Aggregate Location DataApostolos Pyrgelis, Carmela Troncoso, Emiliano De CristofaroNDSS 2018 · 293 citations
Related papers
- Where Have You Been? A Study of Privacy Risk for Point-of-Interest RecommendationKunlin Cai, Jinghuai Zhang, Zhiqing Hong, William Shand et al.KDD 2024 · 5 citations
- TransRisk: Mobility Privacy Risk Prediction based on Transferred KnowledgeXiaoyang Xie, Zhiqing Hong, Zhou Qin, Zhihan Fang et al.UbiComp 2022 · 1 citation
- COLA: Cross-city Mobility Transformer for Human Trajectory SimulationYu Wang, Tongya Zheng, Yuxuan Liang, Shunyu Liu et al.WWW 2024 · 37 citations
- Going Where, by Whom, and at What Time: Next Location Prediction Considering User Preference and Temporal RegularityTianao Sun, Ke Fu, Weiming Huang, Kai Zhao et al.KDD 2024 · 11 citations
- AttnMove: History Enhanced Trajectory Recovery via Attentional NetworkTong Xia, Yunhan Qi, Jie Feng, Fengli Xu et al.AAAI 2021 · 74 citations
