Learning No-Regret Sparse Generalized Linear Models with Varying Observation(s)
Diyang Li, Charles Ling, Zhiqiang Xu, Huan Xiong, Bin Gu
Abstract
Generalized Linear Models (GLMs) encompass a wide array of regression and classification models, where prediction is a function of a linear combination of the input variables. Often in real-world scenarios, a number of observations would be added into or removed from the existing training dataset, necessitating the development of learning systems that can efficiently train optimal models with varying observations in an online (sequential) manner instead of retraining from scratch. Despite the significance of data-varying scenarios, most existing approaches to sparse GLMs concentrate on offline batch updates, leaving online solutions largely underexplored. In this work, we present the first algorithm without compromising accuracy for GLMs regularized by sparsity-enforcing penalties trained on varying observations. Our methodology is capable of handling the addition and deletion of observations simultaneously, while adaptively updating data-dependent regularization parameters to ensure the best statistical performance. Specifically, we recast sparse GLMs as a bilevel optimization objective upon varying observations and characterize it as an explicit gradient flow in the underlying space for the inner and outer subproblems we are optimizing over, respectively. We further derive a set of rules to ensure a proper transition at regions of non-smoothness, and establish the guarantees of theoretical consistency and finite convergence. Encouraging results are exhibited on real-world benchmarks.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 385bbc40-c2f1-4a87-b2a9-1a2ed7e84d3aBuilds on6
- Machine UnlearningLucas Bourtoule, Varun Chandrasekaran, Christopher A. Choquette-Choo, Hengrui Jia et al.S&P 2021 · 1,381 citations
- Manipulating Machine Learning: Poisoning Attacks and Countermeasures for Regression LearningMatthew Jagielski, Alina Oprea, Battista Biggio, Chang Liu et al.S&P 2018 · 867 citations
- Adaptive Machine UnlearningVarun Gupta, Christopher Jung, Seth Neel, Aaron Roth et al.NeurIPS 2021 · 262 citations
- Generalization Error of Generalized Linear Models in High DimensionsMelikasadat Emami, Mojtaba Sahraee-Ardakan, Parthe Pandit, Sundeep Rangan et al.ICML 2020 · 40 citations
- Differentially Private Bayesian Inference for Generalized Linear ModelsTejas D. Kulkarni, Joonas Jälkö, Antti Koskela, Samuel Kaski et al.ICML 2021 · 30 citations
Related papers
- When Online Learning Meets ODE: Learning without Forgetting on Variable Feature SpaceDiyang Li, Bin GuAAAI 2023 · 4 citations
- Chunk Dynamic Updating for Group Lasso with ODEsDiyang Li, Bin GuAAAI 2022 · 2 citations
- Beyond L1: Faster and Better Sparse Models with skglmQuentin Bertrand, Quentin Klopfenstein, Pierre-Antoine Bannier, Gauthier Gidel et al.NeurIPS 2022 · 32 citations
- Towards Fair Disentangled Online Learning for Changing EnvironmentsChen Zhao, Feng Mi, Xintao Wu, Kai Jiang et al.KDD 2023 · 12 citations
- Optimization-Derived Learning with Essential Convergence Analysis of Training and Hyper-trainingRisheng Liu, Xuan Liu, Shangzhi Zeng, Jin Zhang et al.ICML 2022 · 8 citations
