Learning Proximal Operators to Discover Multiple Optima
Lingxiao Li, Noam Aigerman, Vladimir G. Kim, Jiajin Li, Kristjan H. Greenewald, Mikhail Yurochkin, Justin Solomon
摘要
Finding multiple solutions of non-convex optimization problems is a ubiquitous yet challenging task. Most past algorithms either apply single-solution optimization methods from multiple random initial guesses or search in the vicinity of found solutions using ad hoc heuristics. We present an end-to-end method to learn the proximal operator of a family of training problems so that multiple local minima can be quickly obtained from initial guesses by iterating the learned operator, emulating the proximal-point algorithm that has fast convergence. The learned proximal operator can be further generalized to recover multiple optima for unseen problems at test time, enabling applications such as object detection. The key ingredient in our formulation is a proximal regularization term, which elevates the convexity of our training loss: by applying recent theoretical results, we show that for weakly-convex objectives with Lipschitz gradients, training of the proximal operator converges globally with a practical degree of over-parameterization. We further present an exhaustive benchmark for multi-solution optimization to demonstrate the effectiveness of our method.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了最后一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper2
- Towards Constituting Mathematical Structures for Learning to OptimizeJialin Liu, Xiaohan Chen, Zhangyang Wang, Wotao Yin 等ICML 2023 · 被引用 18 次
- Beyond Scores: Proximal Diffusion ModelsZhenghan Fang, Mateo Díaz, Sam Buchanan, Jeremias SulamNeurIPS 2025 · 被引用 6 次
它引用的顶会 Paper6
- Swin Transformer: Hierarchical Vision Transformer using Shifted WindowsZe Liu, Yutong Lin, Yue Cao, Han Hu 等ICCV 2021 · 被引用 31,683 次
- Implicit Neural Representations with Periodic Activation FunctionsVincent Sitzmann, Julien N. P. Martel, Alexander W. Bergman, David B. Lindell 等NeurIPS 2020 · 被引用 4,008 次
- Large-Scale Wasserstein Gradient FlowsPetr Mokrov, Alexander Korotin, Lingxiao Li, Aude Genevay 等NeurIPS 2021 · 被引用 112 次
- CvxNet: Learnable Convex DecompositionBoyang Deng, Kyle Genova, Soroosh Yazdani, Sofien Bouaziz 等CVPR 2020
- NeRD: Neural 3D Reflection Symmetry DetectorYichao Zhou, Shichen Liu, Yi MaCVPR 2021
相关 Paper
- What's in a Prior? Learned Proximal Networks for Inverse ProblemsZhenghan Fang, Sam Buchanan, Jeremias SulamICLR 2024 · 被引用 27 次
- Learn2Hop: Learned Optimization on Rough LandscapesAmil Merchant, Luke Metz, Samuel S. Schoenholz, Ekin D. CubukICML 2021 · 被引用 19 次
- Effective Meta-Regularization by Kernelized Proximal RegularizationWeisen Jiang, James T. Kwok, Yu ZhangNeurIPS 2021 · 被引用 9 次
- Newton Informed Neural Operator for Solving Nonlinear Partial Differential EquationsWenrui Hao, Xinliang Liu, Yahong YangNeurIPS 2024 · 被引用 22 次
- How to Fill the Optimum Set? Population Gradient Descent with Harmless DiversityChengyue Gong, Lemeng Wu, Qiang LiuICML 2022 · 被引用 4 次
