Bayes beats Cross Validation: Efficient and Accurate Ridge Regression via Expectation Maximization
Shu Yu Tew, Mario Boley, Daniel F. Schmidt
Abstract
We present a novel method for tuning the regularization hyper-parameter, , of a ridge regression that is faster to compute than leave-one-out cross-validation (LOOCV) while yielding estimates of the regression parameters of equal, or particularly in the setting of sparse covariates, superior quality to those obtained by minimising the LOOCV risk. The LOOCV risk can suffer from multiple and bad local minima for finite and thus requires the specification of a set of candidate , which can fail to provide good solutions. In contrast, we show that the proposed method is guaranteed to find a unique optimal solution for large enough , under relatively mild conditions, without requiring the specification of any difficult to determine hyper-parameters. This is based on a Bayesian formulation of ridge regression that we prove to have a unimodal posterior for large enough , allowing for both the optimal and the regression coefficients to be jointly learned within an iterative expectation maximization (EM) procedure. Importantly, we show that by utilizing an appropriate preprocessing step, a single iteration of the main EM loop can be implemented in operations, for input data with rows and columns. In contrast, evaluating a single value of using fast LOOCV costs operations when using the same preprocessing. This advantage amounts to an asymptotic improvement of a factor of for candidate values for (in the regime where is the number of regression targets).
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 1db9a362-2fab-4b74-a60e-36bc5e1fb2c9Cited by top-tier papers1
Ask how each one uses itBuilds on1
Related papers
- Ridge Regression: Structure, Cross-Validation, and SketchingSifan Liu, Edgar DobribanICLR 2020 · 52 citations
- Provably tuning the ElasticNet across instancesMaria-Florina Balcan, Misha Khodak, Dravyansh Sharma, Ameet TalwalkarNeurIPS 2022 · 28 citations
- Generalized equivalences between subsampling and ridge regularizationPratik Patil, Jin-Hong DuNeurIPS 2023 · 10 citations
- OKRidge: Scalable Optimal k-Sparse Ridge RegressionJiachang Liu, Sam Rosen, Chudi Zhong, Cynthia RudinNeurIPS 2023 · 10 citations
- Sketching Algorithms and Lower Bounds for Ridge RegressionPraneeth Kacham, David P. WoodruffICML 2022 · 6 citations
