Time-Varying Propensity Score to Bridge the Gap between the Past and Present
Rasool Fakoor, Jonas Mueller, Zachary Chase Lipton, Pratik Chaudhari, Alex Smola
Abstract
Real-world deployment of machine learning models is challenging because data evolves over time. While no model can work when data evolves in an arbitrary fashion, if there is some pattern to these changes, we might be able to design methods to address it. This paper addresses situations when data evolves gradually. We introduce a time-varying propensity score that can detect gradual shifts in the distribution of data which allows us to selectively sample past data to update the model -- not just similar data from the past like that of a standard propensity score but also data that evolved in a similar fashion in the past. The time-varying propensity score is quite general: we demonstrate different ways of implementing it and evaluate it on a variety of problems ranging from supervised learning (e.g., image classification problems) where data undergoes a sequence of gradual shifts, to reinforcement learning tasks (e.g., robotic manipulation and continuous control) where data shifts as the policy or the task changes.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 5a394e4d-cbd8-4f39-ac1b-cfd54d90a1dcCited by top-tier papers2
- Adapting to Continuous Covariate Shift via Online Density Ratio EstimationYu-Jie Zhang, Zhen-Yu Zhang, Peng Zhao, Masashi SugiyamaNeurIPS 2023 · 25 citations
- Prospective Learning: Learning for a Dynamic FutureAshwin De Silva, Rahul Ramesh, Rubing Yang, Siyu Yu et al.NeurIPS 2024 · 5 citations
Builds on10
- Task2Vec: Task Embedding for Meta-LearningAlessandro Achille, Michael Lam, Rahul Tewari, Avinash Ravichandran et al.ICCV 2019 · 359 citations
- A Unified View of Label Shift EstimationSaurabh Garg, Yifan Wu, Sivaraman Balakrishnan, Zachary C. LiptonNeurIPS 2020 · 186 citations
- Meta-Q-LearningRasool Fakoor, Pratik Chaudhari, Stefano Soatto, Alexander J. SmolaICLR 2020 · 162 citations
- Recurrent Model-Free RL Can Be a Strong Baseline for Many POMDPsTianwei Ni, Benjamin Eysenbach, Ruslan SalakhutdinovICML 2022 · 162 citations
- Maximum Likelihood with Bias-Corrected Calibration is Hard-To-Beat at Label Shift AdaptationAmr Alexandari, Anshul Kundaje, Avanti ShrikumarICML 2020 · 123 citations
Related papers
- When to retrain a machine learning modelFlorence Regol, Leo Schwinn, Kyle Sprague, Mark Coates et al.ICML 2025
- Training for the Future: A Simple Gradient Interpolation Loss to Generalize Along TimeAnshul Nasery, Soumyadeep Thakur, Vihari Piratla, Abir De et al.NeurIPS 2021 · 42 citations
- Predictor-corrector algorithms for stochastic optimization under gradual distribution shiftSubha Maity, Debarghya Mukherjee, Moulinath Banerjee, Yuekai SunICLR 2023 · 1 citation
- Evolving Standardization for Continual Domain Generalization over Temporal DriftMixue Xie, Shuang Li, Longhui Yuan, Chi Harold Liu et al.NeurIPS 2023 · 21 citations
- CODA: Temporal Domain Generalization via Concept Drift SimulatorChia-Yuan Chang, Yu-Neng Chuang, Zhimeng Jiang, Kwei-Herng Lai et al.KDD 2025
