Mapping the Multiverse of Latent Representations
Jeremy Wayland, Corinna Coupette, Bastian Rieck
摘要
Echoing recent calls to counter reliability and robustness concerns in machine learning via multiverse analysis, we present PRESTO, a principled framework for mapping the multiverse of machine-learning models that rely on latent representations. Although such models enjoy widespread adoption, the variability in their embeddings remains poorly understood, resulting in unnecessary complexity and untrustworthy representations. Our framework uses persistent homology to characterize the latent spaces arising from different combinations of diverse machine-learning methods, (hyper)parameter configurations, and datasets, allowing us to measure their pairwise (dis)similarity and statistically reason about their distributions. As we demonstrate both theoretically and empirically, our pipeline preserves desirable properties of collections of latent representations, and it can be leveraged to perform sensitivity analysis, detect anomalous embeddings, or efficiently and effectively navigate hyperparameter search spaces.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper5
- Metric Space Magnitude for Evaluating the Diversity of Latent RepresentationsKatharina Limbeck, Rayna Andreeva, Rik Sarkar, Bastian RieckNeurIPS 2024 · 被引用 27 次
- From Bricks to Bridges: Product of Invariances to Enhance Latent Space CommunicationIrene Cannistraci, Luca Moschella, Marco Fumero, Valentino Maiorca 等ICLR 2024 · 被引用 22 次
- No Metric to Rule Them All: Toward Principled Evaluations of Graph-Learning DatasetsCorinna Coupette, Jeremy Wayland, Emily Simons, Bastian RieckICML 2025
- PERSISTENCE SPHERES: BI-CONTINUOUS REPRESENTATIONS OF PERSISTENCE DIAGRAMS.Matteo PegoraroICLR 2026
- Diss-l-ECT: Dissecting Graph Data with Local Euler Characteristic TransformsJulius von Rohrscheidt, Bastian RieckICML 2025
它引用的顶会 Paper19
- Topological AutoencodersMichael Moor, Max Horn, Bastian Rieck, Karsten M. BorgwardtICML 2020 · 被引用 192 次
- Representation Topology Divergence: A Method for Comparing Neural Network RepresentationsSerguei Barannikov, Ilya Trofimov, Nikita Balabin, Evgeny BurnaevICML 2022 · 被引用 69 次
- Optimizer Benchmarking Needs to Account for Hyperparameter TuningPrabhu Teja Sivaprasad, Florian Mai, Thijs Vogels, Martin Jaggi 等ICML 2020 · 被引用 60 次
- On Implicit Regularization in β-VAEsAbhishek Kumar, Ben PooleICML 2020 · 被引用 59 次
- The Shape of Data: Intrinsic Distance for Data DistributionsAnton Tsitsulin, Marina Munkhoeva, Davide Mottin, Panagiotis Karras 等ICLR 2020 · 被引用 57 次
相关 Paper
- Modeling the Machine Learning MultiverseSamuel J. Bell, Onno Kampman, Jesse Dodge, Neil D. LawrenceNeurIPS 2022 · 被引用 23 次
- Milliways: Taming Multiverses through Principled Evaluation of Data Analysis PathsAbhraneel Sarma, Kyle Hwang, Jessica Hullman, Matthew KayCHI 2024 · 被引用 17 次
- Bayesian Optimization for Simultaneous Selection of Machine Learning Algorithms and Hyperparameters on Shared Latent SpaceKazuki Ishikawa, Ryota Ozaki, Yohei Kanzaki, Ichiro Takeuchi 等KDD 2025 · 被引用 1 次
- multiverse: Multiplexing Alternative Data Analyses in R NotebooksAbhraneel Sarma, Alex Kale, Michael Jongho Moon, Nathan Taback 等CHI 2023 · 被引用 22 次
- Improving Self-supervised Molecular Representation Learning using Persistent HomologyYuankai Luo, Lei Shi, Veronika ThostNeurIPS 2023 · 被引用 13 次
