GLOBEM: Cross-Dataset Generalization of Longitudinal Human Behavior Modeling
Xuhai Xu, Xin Liu, Han Zhang, Weichen Wang, Subigya Nepal, Yasaman S. Sefidgar, Woosuk Seo, Kevin S. Kuehn, Jeremy F. Huckins, Margaret E. Morris, Paula S. Nurius, Eve A. Riskin
Abstract
There is a growing body of research revealing that longitudinal passive sensing data from smartphones and wearable devices can capture daily behavior signals for human behavior modeling, such as depression detection. Most prior studies build and evaluate machine learning models using data collected from a single population. However, to ensure that a behavior model can work for a larger group of users, its generalizability needs to be verified on multiple datasets from different populations. We present the first work evaluating cross-dataset generalizability of longitudinal behavior models, using depression detection as an application. We collect multiple longitudinal passive mobile sensing datasets with over 500 users from two institutes over a two-year span, leading to four institute-year datasets. Using the datasets, we closely re-implement and evaluated nine prior depression detection algorithms. Our experiment reveals the lack of model generalizability of these methods. We also implement eight recently popular domain generalization algorithms from the machine learning community. Our results indicate that these methods also do not generalize well on our datasets, with barely any advantage over the naive baseline of guessing the majority. We then present two new algorithms with better generalizability. Our new algorithm, Reorder, significantly and consistently outperforms existing methods on most cross-dataset generalization setups. However, the overall advantage is incremental and still has great room for improvement. Our analysis reveals that the individual differences (both within and between populations) may play the most important role in the cross-dataset generalization challenge. Finally, we provide an open-source benchmark platform GLOBEM- short for Generalization of Longitudinal BEhavior Modeling - to consolidate all 19 algorithms. GLOBEM can support researchers in using, developing, and evaluating different longitudinal behavior modeling methods. We call for researchers' attention to model generalizability evaluation for future longitudinal human behavior modeling studies.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 4fe1958c-7c4f-41ff-8497-2424e1684b01Cited by top-tier papers30
- Mental-LLM: Leveraging Large Language Models for Mental Health Prediction via Online Text DataXuhai Xu, Bingsheng Yao, Yuanzhe Dong, Saadia Gabriel et al.UbiComp 2024 · 281 citations
- Rethinking Human-AI Collaboration in Complex Medical Decision Making: A Case Study in Sepsis DiagnosisShao Zhang, Jianing Yu, Xuhai Xu, Changchang Yin et al.CHI 2024 · 95 citations
- Time2Stop: Adaptive and Explainable Human-AI Loop for Smartphone Overuse InterventionAdiba Orzikulova, Han Xiao, Zhipeng Li, Yukang Yan et al.CHI 2024 · 53 citations
- From Classification to Clinical Insights: Towards Analyzing and Reasoning About Mobile and Behavioral Health Data With Large Language ModelsZachary Englhardt, Chengqian Ma, Margaret E. Morris, Chun-Cheng Chang et al.UbiComp 2024 · 51 citations
- Past, Present, and Future of Sensor-based Human Activity Recognition Using Wearables: A Surveying Tutorial on a Still Challenging TaskHarish Haresamudram, Chi Ian Tang, Sungho Suh, Paul Lukowicz et al.UbiComp 2025 · 31 citations
Builds on14
- WILDS: A Benchmark of in-the-Wild Distribution ShiftsPang Wei Koh, Shiori Sagawa, Henrik Marklund, Sang Michael Xie et al.ICML 2021 · 1,773 citations
- Distributionally Robust Neural NetworksShiori Sagawa, Pang Wei Koh, Tatsunori B. Hashimoto, Percy LiangICLR 2020 · 1,578 citations
- SelfReg: Self-supervised Contrastive Regularization for Domain GeneralizationDaehee Kim, Youngjun Yoo, Seunghyun Park, Jinkyu Kim et al.ICCV 2021 · 338 citations
- Efficient Domain Generalization via Common-Specific Low-Rank DecompositionVihari Piratla, Praneeth Netrapalli, Sunita SarawagiICML 2020 · 250 citations
- Learning explanations that are hard to varyGiambattista Parascandolo, Alexander Neitz, Antonio Orvieto, Luigi Gresele et al.ICLR 2021 · 221 citations
Related papers
- Leveraging Collaborative-Filtering for Personalized Behavior Modeling: A Case Study of Depression Detection among College StudentsXuhai Xu, Prerna Chikersal, Janine M. Dutcher, Yasaman S. Sefidgar et al.UbiComp 2021 · 75 citations
- Generalization and Personalization of Mobile Sensing-Based Mood Inference Models: An Analysis of College Students in Eight CountriesLakmal Meegahapola, William Droz, Peter Kun, Amalia de Götzen et al.UbiComp 2023 · 55 citations
- Predicting Symptom Improvement During Depression Treatment Using Sleep Sensory DataChinmaey Shende, Soumyashree Sahoo, Stephen Sam, Parit Patel et al.UbiComp 2023 · 6 citations
- CrossShift: Quantifying Interpersonal Differences in Mobile Sensing for Mental HealthPanyu Zhang, Minseo Park, Tomiris Ismatzoda, Azizbek Mustafakulov et al.UbiComp 2026
- Biobehavioral Rhythms in Everyday Life: Data and Models for Capturing Cyclic Behavior in Naturalistic SettingsChong Zhao, Maria Ana Cardei, Matthew Clark, Runze Yan et al.UbiComp 2026 · 1 citation
