Learning from Uncertain Data: From Possible Worlds to Possible Models
Jiongli Zhu, Su Feng, Boris Glavic, Babak Salimi
Abstract
We introduce an efficient method for learning linear models from uncertain data, where uncertainty is represented as a set of possible variations in the data, leading to predictive multiplicity. Our approach leverages abstract interpretation and zonotopes, a type of convex polytope, to compactly represent these dataset variations, enabling the symbolic execution of gradient descent on all possible worlds simultaneously. We develop techniques to ensure that this process converges to a fixed point and derive closed-form solutions for this fixed point. Our method provides sound over-approximations of all possible optimal models and viable prediction ranges. We demonstrate the effectiveness of our approach through theoretical and empirical analysis, highlighting its potential to reason about model and prediction uncertainty due to data quality issues in training data.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 2318458a-4eb7-4329-8a4c-ccdfd8a889acBuilds on16
- AI2: Safety and Robustness Certification of Neural Networks with Abstract InterpretationTimon Gehr, Matthew Mirman, Dana Drachsler-Cohen, Petar Tsankov et al.S&P 2018 · 987 citations
- Manipulating Machine Learning: Poisoning Attacks and Countermeasures for Regression LearningMatthew Jagielski, Alina Oprea, Battista Biggio, Chang Liu et al.S&P 2018 · 867 citations
- Predictive Multiplicity in ClassificationCharles T. Marx, Flávio P. Calmon, Berk UstunICML 2020 · 197 citations
- Certified Robustness to Label-Flipping Attacks via Randomized SmoothingElan Rosenfeld, Ezra Winston, Pradeep Ravikumar, J. Zico KolterICML 2020 · 182 citations
- Intrinsic Certified Robustness of Bagging against Data Poisoning AttacksJinyuan Jia, Xiaoyu Cao, Neil Zhenqiang GongAAAI 2021 · 155 citations
Related papers
- Understanding Fixed Predictions via Confined RegionsConnor Lawless, Tsui-Wei Weng, Berk Ustun, Madeleine UdellICML 2025
- Abstract Interpretation of Fixpoint Iterators with Applications to Neural NetworksMark Niklas Müller, Marc Fischer, Robin Staab, Martin T. VechevPLDI 2023 · 5 citations
- A Combinatorial Perspective on the Optimization of Shallow ReLU NetworksMichael Matena, Colin RaffelNeurIPS 2022 · 3 citations
- Zonotope Domains for Lagrangian Neural Network VerificationMatt Jordan, Jonathan Hayase, Alex Dimakis, Sewoong OhNeurIPS 2022 · 6 citations
- Non-parametric Models for Non-negative FunctionsUlysse Marteau-Ferey, Francis R. Bach, Alessandro RudiNeurIPS 2020 · 65 citations
