Bayesian Nonparametrics Meets Data-Driven Distributionally Robust Optimization
Nicola Bariletto, Nhat Ho
Abstract
Training machine learning and statistical models often involves optimizing a data-driven risk criterion. The risk is usually computed with respect to the empirical data distribution, but this may result in poor and unstable out-of-sample performance due to distributional uncertainty. In the spirit of distributionally robust optimization, we propose a novel robust criterion by combining insights from Bayesian nonparametric (i.e., Dirichlet process) theory and a recent decision-theoretic model of smooth ambiguity-averse preferences. First, we highlight novel connections with standard regularized empirical risk minimization techniques, among which Ridge and LASSO regressions. Then, we theoretically demonstrate the existence of favorable finite-sample and asymptotic statistical guarantees on the performance of the robust optimization procedure. For practical implementation, we propose and study tractable approximations of the criterion based on well-known Dirichlet process representations. We also show that the smoothness of the criterion naturally leads to standard gradient-based numerical optimization. Finally, we provide insights into the workings of our method by applying it to a variety of tasks based on simulated and real datasets.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 3084c281-2cde-44e1-8e23-a0b5bcac15f8Cited by top-tier papers1
Ask how each one uses itRelated papers
- Decision Making under the Exponential Family: Distributionally Robust Optimisation with Bayesian Ambiguity SetsCharita Dellaporta, Patrick O'Hara, Theodoros DamoulasICML 2025
- End-to-End Learning for Stochastic Optimization: A Bayesian PerspectiveYves Rychener, Daniel Kuhn, Tobias SutterICML 2023 · 14 citations
- Distributed Distributionally Robust Optimization with Non-Convex ObjectivesYang Jiao, Kai Yang, Dongjin SongNeurIPS 2022 · 21 citations
- Statistical Properties of Robust SatisficingZhiyi Li, Yunbei Xu, Ruohan ZhanICML 2024
- Distributionally Robust Active Learning for Gaussian Process RegressionShion Takeno, Yoshito Okura, Yu Inatsu, Tatsuya Aoyama et al.ICML 2025
