Nuclear Norm Regularization for Deep Learning
Christopher Scarvelis, Justin M. Solomon
摘要
Penalizing the nuclear norm of a function's Jacobian encourages it to locally behave like a low-rank linear map. Such functions vary locally along only a handful of directions, making the Jacobian nuclear norm a natural regularizer for machine learning problems. However, this regularizer is intractable for high-dimensional problems, as it requires computing a large Jacobian matrix and taking its singular value decomposition. We show how to efficiently penalize the Jacobian nuclear norm using techniques tailor-made for deep learning. We prove that for functions parametrized as compositions , one may equivalently penalize the average squared Frobenius norm of and . We then propose a denoising-style approximation that avoids the Jacobian computations altogether. Our method is simple, efficient, and accurate, enabling Jacobian nuclear norm regularization to scale to high-dimensional deep learning problems. We complement our theory with an empirical study of our regularizer's performance and investigate applications to denoising and representation learning.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper2
- Why Deep Jacobian Spectra Separate: Depth-Induced Scaling and Singular-Vector AlignmentNathanaël Haas, François Gatine, Augustin Cosse, Zied BouraouiICML 2026 · 被引用 1 次
- On the Local Complexity of Linear Regions in Deep ReLU NetworksNiket Patel, Guido MontúfarICML 2025
它引用的顶会 Paper4
- High-Resolution Image Synthesis with Latent Diffusion ModelsRobin Rombach, Andreas Blattmann, Dominik Lorenz, Patrick Esser 等CVPR 2022 · 被引用 13,123 次
- Fourier Features Let Networks Learn High Frequency Functions in Low Dimensional DomainsMatthew Tancik, Pratul P. Srinivasan, Ben Mildenhall, Sara Fridovich-Keil 等NeurIPS 2020 · 被引用 4,036 次
- Generalization in diffusion models arises from geometry-adaptive harmonic representationsZahra Kadkhodaie, Florentin Guth, Eero P. Simoncelli, Stéphane MallatICLR 2024 · 被引用 168 次
- Learning Differential Equations that are Easy to SolveJacob Kelly, Jesse Bettencourt, Matthew J. Johnson, David DuvenaudNeurIPS 2020 · 被引用 134 次
相关 Paper
- Learning Representation from Neural Fisher Kernel with Low-rank ApproximationRuixiang Zhang, Shuangfei Zhai, Etai Littwin, Joshua M. SusskindICLR 2022 · 被引用 5 次
- Operator SVD with Neural Networks via Nested Low-Rank ApproximationJongha Jon Ryu, Xiangxiang Xu, Hasan Sabri Melihcan Erol, Yuheng Bu 等ICML 2024 · 被引用 10 次
- Amortized Eigendecomposition for Neural NetworksTianbo Li, Zekun Shi, Jiaxi Zhao, Min LinNeurIPS 2024 · 被引用 4 次
- SKFAC: Training Neural Networks With Faster Kronecker-Factored Approximate CurvatureZedong Tang, Fenlong Jiang, Maoguo Gong, Hao Li 等CVPR 2021
- Fiedler Regularization: Learning Neural Networks with Graph SparsityEdric Tam, David B. DunsonICML 2020 · 被引用 15 次
