Optimizing Nondecomposable Data Dependent Regularizers via Lagrangian Reparameterization Offers Significant Performance and Efficiency Gains
Sathya N. Ravi, Abhay Venkatesh, Glenn Moo Fung, Vikas Singh
摘要
Data dependent regularization is known to benefit a wide variety of problems in machine learning. Often, these regularizers cannot be easily decomposed into a sum over a finite number of terms, e.g., a sum over individual example-wise terms. The F β measure, Area under the ROC curve (AUCROC) and Precision at a fixed recall (P@R) are some prominent examples that are used in many applications. We find that for most medium to large sized datasets, scalability issues severely limit our ability in leveraging the benefits of such regularizers. Importantly, the key technical impediment despite some recent progress is that, such objectives remain difficult to optimize via backpropapagation procedures. While an efficient general-purpose strategy for this problem still remains elusive, in this paper, we show that for many data-dependent nondecomposable regularizers that are relevant in applications, sizable gains in efficiency are possible with minimal code-level changes; in other words, no specialized tools or numerical schemes are needed. Our procedure involves a reparameterization followed by a partial dualization - this leads to a formulation that has provably cheap projection operators. We present a detailed analysis of runtime and convergence properties of our algorithm. On the experimental side, we show that a direct use of our scheme significantly improves the state of the art IOU measures reported for MSCOCO Stuff segmentation dataset.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper2
- Physarum Powered Differentiable Linear Programming Layers and ApplicationsZihang Meng, Sathya N. Ravi, Vikas SinghAAAI 2021 · 被引用 5 次
- Differentiable Optimization of Generalized Nondecomposable Functions using Linear ProgramsZihang Meng, Lopamudra Mukherjee, Yichao Wu, Vikas Singh 等NeurIPS 2021 · 被引用 1 次
相关 Paper
- Large-scale Optimization of Partial AUC in a Range of False Positive RatesYao Yao, Qihang Lin, Tianbao YangNeurIPS 2022 · 被引用 24 次
- Finite-Sum Coupled Compositional Stochastic Optimization: Theory and ApplicationsBokun Wang, Tianbao YangICML 2022 · 被引用 38 次
- Quadruply Stochastic Gradient Method for Large Scale Nonlinear Semi-Supervised Ordinal Regression AUC OptimizationWanli Shi, Bin Gu, Xiang Li, Heng HuangAAAI 2020 · 被引用 13 次
- Exploring the Algorithm-Dependent Generalization of AUPRC Optimization with List StabilityPeisong Wen, Qianqian Xu, Zhiyong Yang, Yuan He 等NeurIPS 2022 · 被引用 15 次
- DeepHoyer: Learning Sparser Neural Network with Differentiable Scale-Invariant Sparsity MeasuresHuanrui Yang, Wei Wen, Hai LiICLR 2020 · 被引用 109 次
