Modeling the Machine Learning Multiverse
Samuel J. Bell, Onno Kampman, Jesse Dodge, Neil D. Lawrence
摘要
Amid mounting concern about the reliability and credibility of machine learning research, we present a principled framework for making robust and generalizable claims: the multiverse analysis. Our framework builds upon the multiverse analysis (Steegen et al., 2016) introduced in response to psychology's own reproducibility crisis. To efficiently explore high-dimensional and often continuous ML search spaces, we model the multiverse with a Gaussian Process surrogate and apply Bayesian experimental design. Our framework is designed to facilitate drawing robust scientific conclusions about model performance, and thus our approach focuses on exploration rather than conventional optimization. In the first of two case studies, we investigate disputed claims about the relative merit of adaptive optimizers. Second, we synthesize conflicting research on the effect of learning rate on the large batch training generalization gap. For the machine learning community, the multiverse analysis is a simple and effective technique for identifying robust claims, for increasing transparency, and a step toward improved reproducibility.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper3
- Mapping the Multiverse of Latent RepresentationsJeremy Wayland, Corinna Coupette, Bastian RieckICML 2024 · 被引用 10 次
- Preventing Harmful Data Practices by using Participatory Input to Navigate the Machine Learning MultiverseJan Simson, Fiona Draxler, Samuel Mehr, Christoph KernCHI 2025 · 被引用 3 次
- Diss-l-ECT: Dissecting Graph Data with Local Euler Characteristic TransformsJulius von Rohrscheidt, Bastian RieckICML 2025
它引用的顶会 Paper4
- Deep Reinforcement Learning at the Edge of the Statistical PrecipiceRishabh Agarwal, Max Schwarzer, Pablo Samuel Castro, Aaron C. Courville 等NeurIPS 2021 · 被引用 1,067 次
- Descending through a Crowded Valley - Benchmarking Deep Learning OptimizersRobin M. Schmidt, Frank Schneider, Philipp HennigICML 2021 · 被引用 195 次
- Hyperparameter Optimization Is Deceiving Us, and How to Stop ItA. Feder Cooper, Yucheng Lu, Jessica Zosa Forde, Christopher De SaNeurIPS 2021 · 被引用 40 次
- Scientific Credibility of Machine Translation Research: A Meta-Evaluation of 769 PapersBenjamin Marie, Atsushi Fujita, Raphael RubinoACL 2021
相关 Paper
- Understanding and Supporting Debugging Workflows in Multiverse AnalysisKen Gu, Eunice Jun, Tim AlthoffCHI 2023 · 被引用 11 次
- Boba: Authoring and Visualizing Multiverse AnalysesYang Liu, Alex Kale, Tim Althoff, Jeffrey HeerIEEE VIS 2020 · 被引用 79 次
- multiverse: Multiplexing Alternative Data Analyses in R NotebooksAbhraneel Sarma, Alex Kale, Michael Jongho Moon, Nathan Taback 等CHI 2023 · 被引用 22 次
- Towards Inferential Reproducibility of Machine Learning ResearchMichael Hagmann, Philipp Meier, Stefan RiezlerICLR 2023 · 被引用 1 次
- Milliways: Taming Multiverses through Principled Evaluation of Data Analysis PathsAbhraneel Sarma, Kyle Hwang, Jessica Hullman, Matthew KayCHI 2024 · 被引用 17 次
