The Relative Instability of Model Comparison with Cross-validation
Alexandre Bayle, Lucas Janson, Lester Mackey
摘要
Cross-validation (CV) is known to provide asymptotically exact tests and confidence intervals for model improvement but only when the model comparison is relatively stable. Surprisingly, we prove that even simple, individually stable models can generate relatively unstable comparisons, calling into question the validity of CV inference. Specifically, we show that the Lasso and its close cousin, soft-thresholding, generate relatively unstable comparisons and invalid CV inferences, even in the most favorable of learning settings and even when both models are individually stable. These findings highlight the importance of verifying relative stability before deploying CV for model comparison.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
相关 Paper
- Cross-validation Confidence Intervals for Test ErrorPierre Bayle, Alexandre Bayle, Lucas Janson, Lester MackeyNeurIPS 2020 · 被引用 76 次
- Can we globally optimize cross-validation loss? Quasiconvexity in ridge regressionWilliam T. Stephenson, Zachary Frangella, Madeleine Udell, Tamara BroderickNeurIPS 2021 · 被引用 15 次
- Deciphering Lasso-based Classification Through a Large Dimensional Analysis of the Iterative Soft-Thresholding AlgorithmMalik Tiomoko, Ekkehard Schnoor, Mohamed El Amine Seddik, Igor Colin 等ICML 2022 · 被引用 4 次
- Cross-Validation for Longitudinal Datasets with Unstable CorrelationsMeera Krishnamoorthy, Michael W. Sjoding, Jenna WiensKDD 2025
- Is Cross-validation the Gold Standard to Estimate Out-of-sample Model Performance?Garud Iyengar, Henry Lam, Tianyu WangNeurIPS 2024 · 被引用 6 次
