On the benefits of maximum likelihood estimation for Regression and Forecasting
Pranjal Awasthi, Abhimanyu Das, Rajat Sen, Ananda Theertha Suresh
摘要
We advocate for a practical Maximum Likelihood Estimation (MLE) approach towards designing loss functions for regression and forecasting, as an alternative to the typical approach of direct empirical risk minimization on a specific target metric. The MLE approach is better suited to capture inductive biases such as prior domain knowledge in datasets, and can output post-hoc estimators at inference time that can optimize different types of target metrics. We present theoretical results to demonstrate that our approach is competitive with any estimator for the target metric under some general conditions. In two example practical settings, Poisson and Pareto regression, we show that our competitive results can be used to prove that the MLE approach has better excess risk bounds than directly minimizing the target metric. We also demonstrate empirically that our method instantiated with a well-designed general purpose mixture likelihood family can obtain superior performance for a variety of tasks across time-series forecasting and regression datasets with different data distributions.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper4
- A decoder-only foundation model for time-series forecastingAbhimanyu Das, Weihao Kong, Rajat Sen, Yichen ZhouICML 2024 · 被引用 601 次
- Unified Training of Universal Time Series Forecasting TransformersGerald Woo, Chenghao Liu, Akshat Kumar, Caiming Xiong 等ICML 2024 · 被引用 513 次
- Neural Spline Search for Quantile Probabilistic ModelingRuoxi Sun, Chun-Liang Li, Sercan Ö. Arik, Michael W. Dusenberry 等AAAI 2023 · 被引用 5 次
- Estimating Unknown Population Sizes Using the Hypergeometric DistributionLiam Hodgson, Danilo BzdokICML 2024
它引用的顶会 Paper3
- Connecting the Dots: Multivariate Time Series Forecasting with Graph Neural NetworksZonghan Wu, Shirui Pan, Guodong Long, Jing Jiang 等KDD 2020 · 被引用 1,738 次
- N-BEATS: Neural basis expansion analysis for interpretable time series forecastingBoris N. Oreshkin, Dmitri Carpov, Nicolas Chapados, Yoshua BengioICLR 2020 · 被引用 1,550 次
- Efficient First-Order Contextual Bandits: Prediction, Allocation, and Triangular DiscriminationDylan J. Foster, Akshay KrishnamurthyNeurIPS 2021 · 被引用 62 次
相关 Paper
- Empirical Gaussian ProcessesJihao Andreas Lin, Sebastian Ament, Louis Tiao, David Eriksson 等ICML 2026
- Calibration by Distribution Matching: Trainable Kernel Calibration MetricsCharlie Marx, Sofian Zalouk, Stefano ErmonNeurIPS 2023 · 被引用 21 次
- Quantile Risk Control: A Flexible Framework for Bounding the Probability of High-Loss PredictionsJake Snell, Thomas P. Zollo, Zhun Deng, Toniann Pitassi 等ICLR 2023 · 被引用 1 次
- Human Pose Regression with Residual Log-likelihood EstimationJiefeng Li, Siyuan Bian, Ailing Zeng, Can Wang 等ICCV 2021 · 被引用 286 次
- Maximum Likelihood Estimation is All You Need for Well-Specified Covariate ShiftJiawei Ge, Shange Tang, Jianqing Fan, Cong Ma 等ICLR 2024 · 被引用 16 次
