Deep Neural Networks Tend To Extrapolate Predictably
Katie Kang, Amrith Setlur, Claire J. Tomlin, Sergey Levine
摘要
Conventional wisdom suggests that neural network predictions tend to be unpredictable and overconfident when faced with out-of-distribution (OOD) inputs. Our work reassesses this assumption for neural networks with high-dimensional inputs. Rather than extrapolating in arbitrary ways, we observe that neural network predictions often tend towards a constant value as input data becomes increasingly OOD. Moreover, we find that this value often closely approximates the optimal constant solution (OCS), i.e., the prediction that minimizes the average loss over the training data without observing the input. We present results showing this phenomenon across 8 datasets with different distributional shifts (including CIFAR10-C and ImageNet-R, S), different loss functions (cross entropy, MSE, and Gaussian NLL), and different architectures (CNNs and transformers). Furthermore, we present an explanation for this behavior, which we first validate empirically and then study theoretically in a simplified setting involving deep homogeneous networks with ReLU activations. Finally, we show how one can leverage our insights in practice to enable risk-sensitive decision-making in the presence of OOD inputs.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper6
- Saddle-to-Saddle Dynamics Explains A Simplicity Bias Across Neural Network ArchitecturesYedi Zhang, Andrew M. Saxe, Peter E. LathamICLR 2026 · 被引用 15 次
- Just One Layer Norm Guarantees Stable ExtrapolationJuliusz Ziomek, George Whittle, Michael A. OsborneNeurIPS 2025 · 被引用 4 次
- Quantifying Distributional Invariance in Causal Subgraph for IRM-Free Graph GeneralizationYang Qiu, Yixiong Zou, Jun Wang, Wei Liu 等NeurIPS 2025 · 被引用 3 次
- Heads collapse, features stay: Why Replay needs big buffersGiulia Lanzillotta, Damiano Meier, Thomas HofmannICLR 2026 · 被引用 3 次
- Fully Heteroscedastic Count Regression with Deep Double Poisson NetworksSpencer Young, Porter Jenkins, Longchao Da, Jeffrey Dotson 等ICML 2025
它引用的顶会 Paper16
- The Many Faces of Robustness: A Critical Analysis of Out-of-Distribution GeneralizationDan Hendrycks, Steven Basart, Norman Mu, Saurav Kadavath 等ICCV 2021 · 被引用 2,294 次
- WILDS: A Benchmark of in-the-Wild Distribution ShiftsPang Wei Koh, Shiori Sagawa, Henrik Marklund, Sang Michael Xie 等ICML 2021 · 被引用 1,773 次
- Out-of-Distribution Detection with Deep Nearest NeighborsYiyou Sun, Yifei Ming, Xiaojin Zhu, Yixuan LiICML 2022 · 被引用 789 次
- CSI: Novelty Detection via Contrastive Learning on Distributionally Shifted InstancesJihoon Tack, Sangwoo Mo, Jongheon Jeong, Jinwoo ShinNeurIPS 2020 · 被引用 755 次
- Gradient Descent Maximizes the Margin of Homogeneous Neural NetworksKaifeng Lyu, Jian LiICLR 2020 · 被引用 402 次
相关 Paper
- Mitigating Neural Network Overconfidence with Logit NormalizationHongxin Wei, Renchunzi Xie, Hao Cheng, Lei Feng 等ICML 2022 · 被引用 386 次
- Towards neural networks that provably know when they don't knowAlexander Meinke, Matthias HeinICLR 2020 · 被引用 151 次
- Entropy Maximization and Meta Classification for Out-of-Distribution Detection in Semantic SegmentationRobin Chan, Matthias Rottmann, Hanno GottschalkICCV 2021 · 被引用 200 次
- The Risks of Invariant Risk MinimizationElan Rosenfeld, Pradeep Kumar Ravikumar, Andrej RisteskiICLR 2021 · 被引用 356 次
- Robustness via Cross-Domain EnsemblesTeresa Yeo, Oguzhan Fatih Kar, Amir ZamirICCV 2021 · 被引用 30 次
