Regularized linear autoencoders recover the principal components, eventually
Xuchan Bao, James Lucas, Sushant Sachdeva, Roger B. Grosse
Abstract
Our understanding of learning input-output relationships with neural nets has improved rapidly in recent years, but little is known about the convergence of the underlying representations, even in the simple case of linear autoencoders (LAEs). We show that when trained with proper regularization, LAEs can directly learn the optimal representation -- ordered, axis-aligned principal components. We analyze two such regularization schemes: non-uniform regularization and a deterministic variant of nested dropout [Rippel et al, ICML' 2014]. Though both regularization schemes converge to the optimal representation, we show that this convergence is slow due to ill-conditioning that worsens with increasing latent dimension. We show that the inefficiency of learning the optimal representation is not inevitable -- we present a simple modification to the gradient descent update that greatly speeds up convergence empirically.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext a5739c29-ff0e-4fb9-bd64-96ab9a4e1125Cited by top-tier papers9
- High-dimensional Asymptotics of Denoising AutoencodersHugo Cui, Lenka ZdeborováNeurIPS 2023 · 26 citations
- The dynamics of representation learning in shallow, non-linear autoencodersMaria Refinetti, Sebastian GoldtICML 2022 · 25 citations
- Identifiability of Deep Polynomial Neural NetworksKonstantin Usevich, Ricardo Augusto Borsoi, Clara Dérand, Marianne ClauselNeurIPS 2025 · 21 citations
- Anytime Sampling for Autoregressive Models via Ordered AutoencodingYilun Xu, Yang Song, Sahaj Garg, Linyuan Gong et al.ICLR 2021 · 15 citations
- Neural Characteristic Activation Analysis and Geometric Parameterization for ReLU NetworksWenlin Chen, Hong GeNeurIPS 2024 · 5 citations
Builds on2
Related papers
- Implicit Rank-Minimizing AutoencoderLi Jing, Jure Zbontar, Yann LeCunNeurIPS 2020 · 63 citations
- It's Enough: Relaxing Diagonal Constraints in Linear Autoencoders for RecommendationJaewan Moon, Hye-young Kim, Jongwuk LeeSIGIR 2023 · 3 citations
- Autoencoder Image Interpolation by Shaping the Latent SpaceAlon Oring, Zohar Yakhini, Yacov Hel-OrICML 2021 · 41 citations
- Fundamental Limits of Two-layer Autoencoders, and Achieving Them with Gradient MethodsAleksandr Shevchenko, Kevin Kögler, Hamed Hassani, Marco MondelliICML 2023 · 3 citations
- Posterior Collapse of a Linear Latent Variable ModelZihao Wang, Liu ZiyinNeurIPS 2022 · 29 citations
