Evaluating Robustness to Dataset Shift via Parametric Robustness Sets
Nikolaj Thams, Michael Oberst, David A. Sontag
Abstract
We give a method for proactively identifying small, plausible shifts in distribution which lead to large differences in model performance. These shifts are defined via parametric changes in the causal mechanisms of observed variables, where constraints on parameters yield a "robustness set" of plausible distributions and a corresponding worst-case loss over the set. While the loss under an individual parametric shift can be estimated via reweighting techniques such as importance sampling, the resulting worst-case optimization problem is non-convex, and the estimate may suffer from large variance. For small shifts, however, we can construct a local second-order approximation to the loss under shift and cast the problem of finding a worst-case shift as a particular non-convex quadratic optimization problem, for which efficient algorithms are available. We demonstrate that this second-order approximation can be estimated directly for shifts in conditional exponential family models, and we bound the approximation error. We apply our approach to a computer vision task (classifying gender from images), revealing sensitivity to shifts in non-causal attributes.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 415a0547-df8d-4937-a443-e0368a4d270cCited by top-tier papers4
- "Why did the Model Fail?": Attributing Model Performance Changes to Distribution ShiftsHaoran Zhang, Harvineet Singh, Marzyeh Ghassemi, Shalmali JoshiICML 2023 · 37 citations
- Achievable distributional robustness when the robust risk is only partially identifiedJulia Kostin, Nicola Gnecco, Fanny YangNeurIPS 2024 · 6 citations
- A Brain-Inspired Gating Mechanism Unlocks Robust Computation in Spiking Neural NetworksQianyi Bai, Haiteng Wang, Qiang YuICLR 2026 · 2 citations
- Going Beyond Static: Understanding Shifts with Time-Series AttributionJiashuo Liu, Nabeel Seedat, Peng Cui, Mihaela van der SchaarICLR 2025
Builds on7
- Distributionally Robust Neural NetworksShiori Sagawa, Pang Wei Koh, Tatsunori B. Hashimoto, Percy LiangICLR 2020 · 1,578 citations
- Leveraging unlabeled data to predict out-of-distribution performanceSaurabh Garg, Sivaraman Balakrishnan, Zachary Chase Lipton, Behnam Neyshabur et al.ICLR 2022 · 160 citations
- Assessing Generalization of SGD via DisagreementYiding Jiang, Vaishnavh Nagarajan, Christina Baek, J. Zico KolterICLR 2022 · 134 citations
- Mandoline: Model Evaluation under Distribution ShiftMayee F. Chen, Karan Goel, Nimit Sharad Sohoni, Fait Poms et al.ICML 2021 · 84 citations
- Out-of-distribution Generalization in the Presence of Nuisance-Induced Spurious CorrelationsAahlad Manas Puli, Lily H. Zhang, Eric Karl Oermann, Rajesh RanganathICLR 2022 · 54 citations
Related papers
- Out-of-Distribution Generalization via Risk Extrapolation (REx)David Krueger, Ethan Caballero, Jörn-Henrik Jacobsen, Amy Zhang et al.ICML 2021 · 1,163 citations
- Towards Understanding Extrapolation: a Causal LensLingjing Kong, Guangyi Chen, Petar Stojanov, Haoxuan Li et al.NeurIPS 2024 · 7 citations
- Not all distributional shifts are equal: Fine-grained robust conformal inferenceJiahao Ai, Zhimei RenICML 2024 · 15 citations
- Statistical Inference Under Constrained Selection BiasSantiago Cortes-Gomez, Mateo Dulce-Rubio, Carlos Miguel Patiño, Bryan WilderICML 2024
- Distributionally Robust Models with Parametric Likelihood RatiosPaul Michel, Tatsunori Hashimoto, Graham NeubigICLR 2022 · 21 citations
