Cauchy-Schwarz Regularizers
Sueda Taner, Ziyi Wang, Christoph Studer
摘要
We introduce a novel class of regularization functions, called Cauchy–Schwarz (CS) regularizers, which can be designed to induce a wide range of properties in solution vectors of optimization problems. To demonstrate the versatility of CS regularizers, we derive regularization functions that promote discrete-valued vectors, eigenvectors of a given matrix, and orthogonal matrices. The resulting CS regularizers are simple, differentiable, and can be free of spurious stationary points, making them suitable for gradient-based solvers and large-scale optimization problems. In addition, CS regularizers automatically adapt to the appropriate scale, which is, for example, beneficial when discretizing the weights of neural networks. To demonstrate the efficacy of CS regularizers, we provide results for solving underdetermined systems of linear equations and weight quantization in neural networks. Furthermore, we discuss specializations, variations, and generalizations, which lead to an even broader class of new and possibly more powerful regularizers.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper6
- Differentiable Soft Quantization: Bridging Full-Precision and Low-Bit Neural NetworksRuihao Gong, Xianglong Liu, Shenghu Jiang, Tianxiang Li 等ICCV 2019 · 被引用 540 次
- BSQ: Exploring Bit-Level Sparsity for Mixed-Precision Neural Network QuantizationHuanrui Yang, Lin Duan, Yiran Chen, Hai LiICLR 2021 · 被引用 83 次
- Distance-aware QuantizationDohyung Kim, Junghyup Lee, Bumsub HamICCV 2021 · 被引用 42 次
- Improved Imaging by Invex Regularizers with Global Optima GuaranteesSamuel Pinilla, Tingting Mu, Neil Bourne, Jeyan ThiyagalingamNeurIPS 2022 · 被引用 11 次
- Resilient Binary Neural NetworkSheng Xu, Yanjing Li, Teli Ma, Mingbao Lin 等AAAI 2023 · 被引用 1 次
相关 Paper
- Obtaining Adjustable Regularization for Free via Iterate AveragingJingfeng Wu, Vladimir Braverman, Lin YangICML 2020 · 被引用 2 次
- Network Quantization With Element-Wise Gradient ScalingJunghyup Lee, Dohyung Kim, Bumsub HamCVPR 2021
- Deep linear networks for regression are implicitly regularized towards flat minimaPierre Marion, Lénaïc ChizatNeurIPS 2024 · 被引用 21 次
- Deep Sturm-Liouville: From Sample-Based to 1D Regularization with Learnable Orthogonal Basis FunctionsDavid Vigouroux, Joseba Dalmau, Louis Béthune, Victor BoutinICML 2025
- Improving Deep Learning Speed and Performance Through Synaptic Neural BalanceAntonios Alexos, Ian Domingo, Pierre BaldiAAAI 2025
