Lune

ICML2026顶会

TPV: Parameter Perturbations Through the Lens of Test Prediction Variance

Devansh Arpit

2026年份

摘要

We introduce test prediction variance (TPV)—the first-order sensitivity of a trained model's outputs to parameter perturbations—as a unifying framework for analyzing post-training robustness. TPV's trace form Tr(HeffC)\mathrm{Tr}(H_{\mathrm{eff}}C) separates the geometry of the trained model HeffH_{\mathrm{eff}} from the perturbation covariance CC, placing SGD noise, label noise, quantization, and pruning under a single lens. The resulting expressions recover the wide-minima hypothesis for SGD and quantization noise, and yield a distinct Jacobian-spectral characterization for label noise connecting label-noise TPV with benign overfitting in nonlinear networks. Theoretically, we prove that training-set TPV converges to its test-set counterpart in the overparameterized limit, irrespective of generalization performance, providing the first result that prediction variance under local parameter perturbations can be inferred from training inputs alone. Empirically, this stability holds far more broadly, including at very low widths. Further, TPV correlates well with test loss, enabling practical applications: JBR, a label-free pruning criterion derived from TPV geometry matching state-of-the-art baselines; and training-set based model selection signal for in-distribution and transfer learning scenarios. Code Available Here

问问这篇 Paper

智能体会读完全文。

Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。

可以从这些问题问起

智能体调用

Luneget_paper_fulltext

在 Lune 里问

免费开始,无需绑卡

它引用的顶会 Paper9

相关 Paper

黄昏的海面,两侧是细线勾勒的悬崖