On Lottery Tickets and Minimal Task Representations in Deep Reinforcement Learning
Marc Aurel Vischer, Robert Tjarko Lange, Henning Sprekeler
Abstract
The lottery ticket hypothesis questions the role of overparameterization in supervised deep learning. But how is the performance of winning lottery tickets affected by the distributional shift inherent to reinforcement learning problems? In this work, we address this question by comparing sparse agents who have to address the non-stationarity of the exploration-exploitation problem with supervised agents trained to imitate an expert. We show that feed-forward networks trained with behavioural cloning compared to reinforcement learning can be pruned to higher levels of sparsity without performance degradation. This suggests that in order to solve the RL-specific distributional shift agents require more degrees of freedom. Using a set of carefully designed baseline conditions, we find that the majority of the lottery ticket effect in both learning paradigms can be attributed to the identified mask rather than the weight initialization. The input layer mask selectively prunes entire input dimensions that turn out to be irrelevant for the task at hand. At a moderate level of sparsity the mask identified by iterative magnitude pruning yields minimal task-relevant representations, i.e., an interpretable inductive bias. Finally, we propose a simple initialization rescaling which promotes the robust identification of sparse task representations in low-dimensional control tasks.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext e6bb56d1-f367-4571-a0cd-83b9a5e1624dCited by top-tier papers11
- The State of Sparse Training in Deep Reinforcement LearningLaura Graesser, Utku Evci, Erich Elsen, Pablo Samuel CastroICML 2022 · 65 citations
- In value-based deep reinforcement learning, a pruned network is a good networkJohan S. Obando-Ceron, Aaron C. Courville, Pablo Samuel CastroICML 2024 · 36 citations
- Lottery Tickets on a Data Diet: Finding Initializations with Sparse Trainable NetworksMansheej Paul, Brett W. Larsen, Surya Ganguli, Jonathan Frankle et al.NeurIPS 2022 · 27 citations
- Adversarial Cheap TalkChris Lu, Timon Willi, Alistair Letcher, Jakob Nicolaus FoersterICML 2023 · 17 citations
- Differentiable Transportation PruningYunqiang Li, Jan C. van Gemert, Torsten Hoefler, Bert Moons et al.ICCV 2023 · 17 citations
Builds on7
- Linear Mode Connectivity and the Lottery Ticket HypothesisJonathan Frankle, Gintare Karolina Dziugaite, Daniel M. Roy, Michael CarbinICML 2020 · 750 citations
- The Lottery Ticket Hypothesis for Pre-trained BERT NetworksTianlong Chen, Jonathan Frankle, Shiyu Chang, Sijia Liu et al.NeurIPS 2020 · 428 citations
- The Early Phase of Neural Network TrainingJonathan Frankle, David J. Schwab, Ari S. MorcosICLR 2020 · 199 citations
- Playing the lottery with rewards and multiple languages: lottery tickets in RL and NLPHaonan Yu, Sergey Edunov, Yuandong Tian, Ari S. MorcosICLR 2020 · 156 citations
- Long Live the Lottery: The Existence of Winning Tickets in Lifelong LearningTianlong Chen, Zhenyu Zhang, Sijia Liu, Shiyu Chang et al.ICLR 2021 · 23 citations
Related papers
- Neural Network Pruning Denoises the Features and Makes Local Connectivity Emerge in Visual TasksFranco Pellegrini, Giulio BiroliICML 2022 · 10 citations
- Analyzing Lottery Ticket Hypothesis from PAC-Bayesian Theory PerspectiveKeitaro Sakamoto, Issei SatoNeurIPS 2022 · 11 citations
- Find A Winning Sign: Sign Is All We Need to Win the LotteryJunghun Oh, Sungyong Baik, Kyoung Mu LeeICLR 2025
- Masks, Signs, And Learning Rate RewindingAdvait Harshal Gadhikar, Rebekka BurkholzICLR 2024 · 15 citations
- Sparse Training from Random Initialization: Aligning Lottery Ticket Masks using Weight SymmetryMohammed Adnan, Rohan Jain, Ekansh Sharma, Rahul Krishnan et al.ICML 2025
