MiniMax Learning of Interpretable Factored Stochastic Policies from Conjoint Data, with Uncertainty Quantification
Connor T Jerzak, Priyanshi Chandra, Rishi Hazra
摘要
We study offline policy optimization over exponentially large factorial action spaces from randomized preference data, showing how conjoint experiments can estimate interpretable stochastic policies with asymptotically valid uncertainty under regularity conditions. Conjoint analyses typically report Average Marginal Component Effects (AMCEs) by averaging over opponent attributes and thus ignore strategic interdependence. We instead learn stochastic interventions—product-of-Categorical policies over factor levels—that (i) optimize expected outcomes in an average-case setting and (ii) extend to a two-player minimax (adversarial) setting that realistically captures simultaneous strategic candidate selection. Methodologically, we derive a closed-form optimizer for a tractable two-way interaction regime with variance regularization, and provide a general gradient-based procedure for richer model classes. Uncertainty from the outcome model propagates asymptotically to both the optimal policy and its value via a Delta method approximation. We further model institutional details (e.g., primaries) inside the minimax objective and introduce a data-driven measure of strategic divergence between parties. On synthetic data, we empirically characterize finite-sample error and coverage as dimensionality and vary. On a U.S. presidential conjoint, adversarially learned policies produce restricted-equilibrium vote shares that align with historical election ranges in our data, in stark contrast to non-adversarial (averaging) optimizers.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
相关 Paper
- A Minimax Approach for Optimal Intervention Policy Learning with Two-Stage OutcomesChenyang Li, Hao Mei, Yue LiuICML 2026
- Synthetic Combinations: A Causal Inference Framework for Combinatorial InterventionsAbhineet Agarwal, Anish Agarwal, Suhas VijaykumarNeurIPS 2023 · 被引用 14 次
- USCO-Solver: Solving Undetermined Stochastic Combinatorial Optimization ProblemsGuangmo TongNeurIPS 2021 · 被引用 7 次
- Interpolated Stochastic Interventions Based on Propensity Scores, Target Policies and Treatment-Specific CostsJohan de AguasAAAI 2026
- One-Shot Strategic Classification Under Unknown CostsElan Rosenfeld, Nir RosenfeldICML 2024 · 被引用 10 次
