Emergence of Sparse Representations from Noise
Trenton Bricken, Rylan Schaeffer, Bruno A. Olshausen, Gabriel Kreiman
摘要
A hallmark of biological neural networks, which distinguishes them from their artificial counterparts, is the high degree of sparsity in their activations. This discrepancy raises three questions our work helps to answer: (i) Why are biological networks so sparse? (ii) What are the benefits of this sparsity? (iii) How can these benefits be utilized by deep learning models? Our answers to all of these questions center around training networks to handle random noise. Surprisingly, we discover that noisy training introduces three implicit loss terms that result in sparsely firing neurons specializing to high variance features of the dataset. When trained to reconstruct noisy-CIFAR10, neurons learn biological receptive fields. More broadly, noisy training presents a new approach to potentially increase model interpretability with additional benefits to robustness and computational efficiency.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper4
- Brain-like Variational InferenceHadi Vafaii, Dekel Galor, Jacob L. YatesNeurIPS 2025 · 被引用 7 次
- On the Relationship Between Activation Outliers and Feature Death in Sparse AutoencodersElana Simon, Etowah Adams, James ZouICML 2026
- Emergence and Effectiveness of Task Vectors in In-Context Learning: An Encoder Decoder PerspectiveSeungwook Han, Jinyeop Song, Jeff Gore, Pulkit AgrawalICML 2025
- Bilinear MLPs enable weight-based mechanistic interpretabilityMichael T. Pearce, Thomas Dooms, Alice Rigg, José Oramas 等ICLR 2025
它引用的顶会 Paper8
- On Adaptive Attacks to Adversarial Example DefensesFlorian Tramèr, Nicholas Carlini, Wieland Brendel, Aleksander MadryNeurIPS 2020 · 被引用 1,026 次
- Explicit Regularisation in Gaussian Noise InjectionsAlexander Camuto, Matthew Willetts, Umut Simsekli, Stephen J. Roberts 等NeurIPS 2020 · 被引用 90 次
- SGD with Large Step Sizes Learns Sparse FeaturesMaksym Andriushchenko, Aditya Vardhan Varre, Loucas Pillaud-Vivien, Nicolas FlammarionICML 2023 · 被引用 77 次
- Powerpropagation: A sparsity inducing weight reparameterisationJonathan Schwarz, Siddhant M. Jayakumar, Razvan Pascanu, Peter E. Latham 等NeurIPS 2021 · 被引用 63 次
- On Implicit Regularization in β-VAEsAbhishek Kumar, Ben PooleICML 2020 · 被引用 59 次
相关 Paper
- Nonlinear dynamics of localization in neural receptive fieldsLeon Lufkin, Andrew M. Saxe, Erin GrantNeurIPS 2024 · 被引用 2 次
- Revisiting Sparse Convolutional Model for Visual RecognitionXili Dai, Mingyang Li, Pengyuan Zhai, Shengbang Tong 等NeurIPS 2022 · 被引用 45 次
- Finding trainable sparse networks through Neural Tangent TransferTianlin Liu, Friedemann ZenkeICML 2020 · 被引用 40 次
- Neural Sparse Representation for Image RestorationYuchen Fan, Jiahui Yu, Yiqun Mei, Yulun Zhang 等NeurIPS 2020 · 被引用 39 次
- The computational and learning benefits of Daleian neural networksAdam Haber, Elad SchneidmanNeurIPS 2022 · 被引用 10 次
