Deep Sturm-Liouville: From Sample-Based to 1D Regularization with Learnable Orthogonal Basis Functions
David Vigouroux, Joseba Dalmau, Louis Béthune, Victor Boutin
Abstract
Although Artificial Neural Networks (ANNs) have achieved remarkable success across various tasks, they still suffer from limited generalization. We hypothesize that this limitation arises from the traditional sample-based (0-dimensionnal) regularization used in ANNs. To overcome this, we introduce Deep Sturm-Liouville (DSL), a novel function approximator that enables continuous 1D regularization along field lines in the input space by integrating the Sturm-Liouville Theorem (SLT) into the deep learning framework. DSL defines field lines traversing the input space, along which a Sturm-Liouville problem is solved to generate orthogonal basis functions, enforcing implicit regularization thanks to the desirable properties of SLT. These basis functions are linearly combined to construct the DSL approximator. Both the vector field and basis functions are parameterized by neural networks and learned jointly. We demonstrate that the DSL formulation naturally arises when solving a Rank-1 Parabolic Eigenvalue Problem. DSL is trained efficiently using stochastic gradient descent via implicit differentiation. DSL achieves competitive performance and demonstrate improved sample efficiency on diverse multivariate datasets including high-dimensional image datasets such as MNIST and CIFAR-10.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Builds on3
- On the distance between two neural networks and the stability of learningJeremy Bernstein, Arash Vahdat, Yisong Yue, Ming-Yu LiuNeurIPS 2020 · 77 citations
- Boosting Sample Efficiency and Generalization in Multi-agent Reinforcement Learning via EquivarianceJoshua McClellan, Naveed Haghani, John Winder, Furong Huang et al.NeurIPS 2024 · 21 citations
- Neural Network Approximations of PDEs Beyond Linearity: A Representational PerspectiveTanya Marwah, Zachary Chase Lipton, Jianfeng Lu, Andrej RisteskiICML 2023 · 16 citations
Related papers
- Functional Regularization for Reinforcement Learning via Learned Fourier FeaturesAlexander C. Li, Deepak PathakNeurIPS 2021 · 27 citations
- ReLUs Are Sufficient for Learning Implicit Neural RepresentationsJoseph Shenouda, Yamin Zhou, Robert D. NowakICML 2024 · 7 citations
- Total Deep Variation for Linear Inverse ProblemsErich Kobler, Alexander Effland, Karl Kunisch, Thomas PockCVPR 2020
- Deep Latent Regularity Network for Modeling Stochastic Partial Differential EquationsShiqi Gong, Peiyan Hu, Qi Meng, Yue Wang et al.AAAI 2023 · 7 citations
- D4FT: A Deep Learning Approach to Kohn-Sham Density Functional TheoryTianbo Li, Min Lin, Zheyuan Hu, Kunhao Zheng et al.ICLR 2023 · 1 citation
