Local Redundancy: An Information-Theoretic Measure of Plasticity from Synthetic Memorization
Jiaxuan Cheng
Abstract
Plasticity—a neural network's ability to adapt to new tasks—is critical for continual and transfer learning. Existing measures, such as effective rank, dead neuron fraction, and weight norm, lack theoretical grounding and correlate poorly with performance on new tasks. We introduce local redundancy , an information-theoretic measure derived from universal compression theory. We define local redundancy as the worst-case redundancy of a local model family—parameters in an infinitesimal neighborhood along gradient directions—and show this is a principled measure of plasticity. Although local redundancy is intractable to compute exactly, we prove that the expected squared gradient norm on a synthetic memorization task provides an efficiently computable lower bound. Experiments on continual image classification and time series transfer learning demonstrate that local redundancy predicts downstream performance better than existing measures and enables pretraining checkpoint selection where validation loss plateaus.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Builds on12
- Searching for MobileNetV3Andrew Howard, Ruoming Pang, Hartwig Adam, Quoc V. Le et al.ICCV 2019 · 9,163 citations
- Informer: Beyond Efficient Transformer for Long Sequence Time-Series ForecastingHaoyi Zhou, Shanghang Zhang, Jieqi Peng, Shuai Zhang et al.AAAI 2021 · 7,289 citations
- Linear Mode Connectivity and the Lottery Ticket HypothesisJonathan Frankle, Gintare Karolina Dziugaite, Daniel M. Roy, Michael CarbinICML 2020 · 750 citations
- A Time Series is Worth 64 Words: Long-term Forecasting with TransformersYuqi Nie, Nam H. Nguyen, Phanwadee Sinthong, Jayant KalagnanamICLR 2023 · 536 citations
- On Warm-Starting Neural Network TrainingJordan T. Ash, Ryan P. AdamsNeurIPS 2020 · 288 citations
Related papers
- PLATE: Plasticity-Tunable Efficient Adapters for Geometry-Aware Continual LearningRomain CosentinoICML 2026
- Auto-Compressing NetworksEvangelos Dorovatas, Georgios Paraskevopoulos, Alexandros PotamianosNeurIPS 2025 · 4 citations
- CPR: Classifier-Projection Regularization for Continual LearningSungmin Cha, Hsiang Hsu, Taebaek Hwang, Flávio P. Calmon et al.ICLR 2021 · 32 citations
- Ranking Neural CheckpointsYandong Li, Xuhui Jia, Ruoxin Sang, Yukun Zhu et al.CVPR 2021
- Self-Normalized Resets for Plasticity in Continual LearningVivek F. Farias, Adam Daniel JozefiakICLR 2025
