MiniMax Learning of Interpretable Factored Stochastic Policies from Conjoint Data, with Uncertainty Quantification
Connor T Jerzak, Priyanshi Chandra, Rishi Hazra
Abstract
We study offline policy optimization over exponentially large factorial action spaces from randomized preference data, showing how conjoint experiments can estimate interpretable stochastic policies with asymptotically valid uncertainty under regularity conditions. Conjoint analyses typically report Average Marginal Component Effects (AMCEs) by averaging over opponent attributes and thus ignore strategic interdependence. We instead learn stochastic interventions—product-of-Categorical policies over factor levels—that (i) optimize expected outcomes in an average-case setting and (ii) extend to a two-player minimax (adversarial) setting that realistically captures simultaneous strategic candidate selection. Methodologically, we derive a closed-form optimizer for a tractable two-way interaction regime with variance regularization, and provide a general gradient-based procedure for richer model classes. Uncertainty from the outcome model propagates asymptotically to both the optimal policy and its value via a Delta method approximation. We further model institutional details (e.g., primaries) inside the minimax objective and introduce a data-driven measure of strategic divergence between parties. On synthetic data, we empirically characterize finite-sample error and coverage as dimensionality and vary. On a U.S. presidential conjoint, adversarially learned policies produce restricted-equilibrium vote shares that align with historical election ranges in our data, in stark contrast to non-adversarial (averaging) optimizers.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 3b063cfa-5653-45ec-8a15-08508ef0093cRelated papers
- A Minimax Approach for Optimal Intervention Policy Learning with Two-Stage OutcomesChenyang Li, Hao Mei, Yue LiuICML 2026
- Synthetic Combinations: A Causal Inference Framework for Combinatorial InterventionsAbhineet Agarwal, Anish Agarwal, Suhas VijaykumarNeurIPS 2023 · 14 citations
- USCO-Solver: Solving Undetermined Stochastic Combinatorial Optimization ProblemsGuangmo TongNeurIPS 2021 · 7 citations
- Interpolated Stochastic Interventions Based on Propensity Scores, Target Policies and Treatment-Specific CostsJohan de AguasAAAI 2026
- One-Shot Strategic Classification Under Unknown CostsElan Rosenfeld, Nir RosenfeldICML 2024 · 10 citations
