Hyperparameter Tuning is All You Need for LISTA
Xiaohan Chen, Jialin Liu, Zhangyang Wang, Wotao Yin
Abstract
Learned Iterative Shrinkage-Thresholding Algorithm (LISTA) introduces the concept of unrolling an iterative algorithm and training it like a neural network. It has had great success on sparse recovery. In this paper, we show that adding momentum to intermediate variables in the LISTA network achieves a better convergence rate and, in particular, the network with instance-optimal parameters is superlinearly convergent. Moreover, our new theoretical results lead to a practical approach of automatically and adaptively calculating the parameters of a LISTA network layer based on its previous layers. Perhaps most surprisingly, such an adaptive-parameter procedure reduces the training of LISTA to tuning only three hyperparameters from data: a new record set in the context of the recent advances on trimming down LISTA complexity. We call this new ultra-light weight network HyperLISTA. Compared to state-of-the-art LISTA models, HyperLISTA achieves almost the same performance on seen data distributions and performs better when tested on unseen distributions (specifically, those with different sparsity levels and nonzero magnitudes). Code is available: https://github.com/VITA-Group/HyperLISTA .
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 7ec4aa27-fa01-40cf-8a9d-e42742f3d234Cited by top-tier papers8
- Symbolic Learning to Optimize: Towards Interpretability and ScalabilityWenqing Zheng, Tianlong Chen, Ting-Kuei Hu, Zhangyang WangICLR 2022 · 21 citations
- Towards Constituting Mathematical Structures for Learning to OptimizeJialin Liu, Xiaohan Chen, Zhangyang Wang, Wotao Yin et al.ICML 2023 · 18 citations
- Deep FlexQP: Accelerated Nonlinear Programming via Deep UnfoldingAlex Oshin, Rahul Vodeb Ghosh, Augustinos D. Saravanos, Evangelos A. TheodorouICLR 2026 · 7 citations
- Unrolled denoising networks provably learn to perform optimal Bayesian inferenceAayush Karan, Kulin Shah, Sitan Chen, Yonina C. EldarNeurIPS 2024 · 5 citations
- Non-Asymptotic Uncertainty Quantification in High-Dimensional LearningFrederik Hoppe, Claudio Mayrink Verdun, Hannah Laus, Felix Krahmer et al.NeurIPS 2024 · 5 citations
Builds on4
- Learned Robust PCA: A Scalable Deep Unfolding Approach for High-Dimensional Outlier DetectionHanQin Cai, Jialin Liu, Wotao YinNeurIPS 2021 · 69 citations
- Sparse Coding with Gated Learned ISTAKailun Wu, Yiwen Guo, Ziang Li, Changshui ZhangICLR 2020 · 43 citations
- A Design Space Study for LISTA and BeyondTianjian Meng, Xiaohan Chen, Yifan Jiang, Zhangyang WangICLR 2021 · 3 citations
- Neurally Augmented ALISTAFreya Behrens, Jonathan Sauder, Peter JungICLR 2021 · 1 citation
Related papers
- Learned Extragradient ISTA with Interpretable Residual Structures for Sparse CodingYangyang Li, Lin Kong, Fanhua Shang, Yuanyuan Liu et al.AAAI 2021 · 13 citations
- A Unified Framework for Soft Threshold PruningYanqi Chen, Zhengyu Ma, Wei Fang, Xiawu Zheng et al.ICLR 2023 · 6 citations
- Provable Learning-based Algorithm For Sparse RecoveryXinshi Chen, Haoran Sun, Le SongICLR 2022
- Bilevel Optimization under Unbounded Smoothness: A New Algorithm and Convergence AnalysisJie Hao, Xiaochuan Gong, Mingrui LiuICLR 2024 · 14 citations
- Fast Hierarchical Deep Unfolding Network for Image Compressed SensingWenxue Cui, Shaohui Liu, Debin ZhaoACM MM 2022 · 15 citations
