Statistical Inference Under Constrained Selection Bias
Santiago Cortes-Gomez, Mateo Dulce-Rubio, Carlos Miguel Patiño, Bryan Wilder
摘要
Large-scale datasets are increasingly being used to inform decision making. While this effort aims to ground policy in real-world evidence, challenges have arisen as selection bias and other forms of distribution shifts often plague observational data. Previous attempts to provide robust inference have given guarantees depending on a user-specified amount of possible distribution shift (e.g., the maximum KL divergence between the observed and target distributions). However, decision makers will often have additional knowledge about the target distribution which constrains the kind of possible shifts. To leverage such information, we propose a framework that enables statistical inference in the presence of selection bias which obeys user-specified constraints in the form of functions whose expectation is known under the target distribution. The output is high-probability bounds on the value of an estimand for the target distribution. Hence, our method leverages domain knowledge in order to partially identify a wide class of estimands. We analyze the computational and statistical properties of methods to estimate these bounds and show that our method can produce informative bounds on a variety of simulated and semisynthetic tasks, as well as in a real-world use case.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper2
- Domain constraints improve risk prediction when outcome data is missingSidhika Balachandar, Nikhil Garg, Emma PiersonICLR 2024 · 被引用 11 次
- Conformal Mixed-Integer Constraint Learning with Feasibility GuaranteesDaniel Ovalle, Lorenz T. Biegler, Ignacio E. Grossmann, Carl D. Laird 等NeurIPS 2025 · 被引用 2 次
它引用的顶会 Paper4
- Retiring Adult: New Datasets for Fair Machine LearningFrances Ding, Moritz Hardt, John Miller, Ludwig SchmidtNeurIPS 2021 · 被引用 671 次
- A Generative Adversarial Framework for Bounding Confounded Causal EffectsYaowei Hu, Yongkai Wu, Lu Zhang, Xintao WuAAAI 2021 · 被引用 32 次
- Partial Identification of Treatment Effects with Implicit Generative ModelsVahid Balazadeh Meresht, Vasilis Syrgkanis, Rahul G. KrishnanNeurIPS 2022 · 被引用 25 次
- The s-value: evaluating stability with respect to distributional shiftsSuyash Gupta, Dominik RothenhäuslerNeurIPS 2023 · 被引用 21 次
相关 Paper
- Robust Generalization despite Distribution Shift via Minimum Discriminating InformationTobias Sutter, Andreas Krause, Daniel KuhnNeurIPS 2021 · 被引用 13 次
- Falsification before Extrapolation in Causal Effect EstimationZeshan M. Hussain, Michael Oberst, Ming-Chieh Shih, David A. SontagNeurIPS 2022 · 被引用 11 次
- What's the Harm? Sharp Bounds on the Fraction Negatively Affected by TreatmentNathan KallusNeurIPS 2022 · 被引用 40 次
- Towards Estimating Bounds on the Effect of Policies under Unobserved ConfoundingAlexis Bellot, Silvia ChiappaNeurIPS 2024 · 被引用 6 次
- Estimation of Bounds on Potential Outcomes For Decision MakingMaggie Makar, Fredrik D. Johansson, John V. Guttag, David A. SontagICML 2020 · 被引用 11 次
