Deep Sturm-Liouville: From Sample-Based to 1D Regularization with Learnable Orthogonal Basis Functions
David Vigouroux, Joseba Dalmau, Louis Béthune, Victor Boutin
摘要
Although Artificial Neural Networks (ANNs) have achieved remarkable success across various tasks, they still suffer from limited generalization. We hypothesize that this limitation arises from the traditional sample-based (0-dimensionnal) regularization used in ANNs. To overcome this, we introduce Deep Sturm-Liouville (DSL), a novel function approximator that enables continuous 1D regularization along field lines in the input space by integrating the Sturm-Liouville Theorem (SLT) into the deep learning framework. DSL defines field lines traversing the input space, along which a Sturm-Liouville problem is solved to generate orthogonal basis functions, enforcing implicit regularization thanks to the desirable properties of SLT. These basis functions are linearly combined to construct the DSL approximator. Both the vector field and basis functions are parameterized by neural networks and learned jointly. We demonstrate that the DSL formulation naturally arises when solving a Rank-1 Parabolic Eigenvalue Problem. DSL is trained efficiently using stochastic gradient descent via implicit differentiation. DSL achieves competitive performance and demonstrate improved sample efficiency on diverse multivariate datasets including high-dimensional image datasets such as MNIST and CIFAR-10.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper3
- On the distance between two neural networks and the stability of learningJeremy Bernstein, Arash Vahdat, Yisong Yue, Ming-Yu LiuNeurIPS 2020 · 被引用 77 次
- Boosting Sample Efficiency and Generalization in Multi-agent Reinforcement Learning via EquivarianceJoshua McClellan, Naveed Haghani, John Winder, Furong Huang 等NeurIPS 2024 · 被引用 21 次
- Neural Network Approximations of PDEs Beyond Linearity: A Representational PerspectiveTanya Marwah, Zachary Chase Lipton, Jianfeng Lu, Andrej RisteskiICML 2023 · 被引用 16 次
相关 Paper
- Functional Regularization for Reinforcement Learning via Learned Fourier FeaturesAlexander C. Li, Deepak PathakNeurIPS 2021 · 被引用 27 次
- ReLUs Are Sufficient for Learning Implicit Neural RepresentationsJoseph Shenouda, Yamin Zhou, Robert D. NowakICML 2024 · 被引用 7 次
- Total Deep Variation for Linear Inverse ProblemsErich Kobler, Alexander Effland, Karl Kunisch, Thomas PockCVPR 2020
- Deep Latent Regularity Network for Modeling Stochastic Partial Differential EquationsShiqi Gong, Peiyan Hu, Qi Meng, Yue Wang 等AAAI 2023 · 被引用 7 次
- D4FT: A Deep Learning Approach to Kohn-Sham Density Functional TheoryTianbo Li, Min Lin, Zheyuan Hu, Kunhao Zheng 等ICLR 2023 · 被引用 1 次
