Symbolic Learning to Optimize: Towards Interpretability and Scalability
Wenqing Zheng, Tianlong Chen, Ting-Kuei Hu, Zhangyang Wang
摘要
Recent studies on Learning to Optimize (L2O) suggest a promising path to automating and accelerating the optimization procedure for complicated tasks. Existing L2O models parameterize optimization rules by neural networks, and learn those numerical rules via meta-training. However, they face two common pitfalls: (1) scalability: the numerical rules represented by neural networks create extra memory overhead for applying L2O models, and limit their applicability to optimizing larger tasks; (2) interpretability: it is unclear what an L2O model has learned in its black-box optimization rule, nor is it straightforward to compare different L2O models in an explainable way. To avoid both pitfalls, this paper proves the concept that we can "kill two birds by one stone", by introducing the powerful tool of symbolic regression to L2O. In this paper, we establish a holistic symbolic representation and analysis framework for L2O, which yields a series of insights for learnable optimizers. Leveraging our findings, we further propose a lightweight L2O model that can be meta-trained on large-scale problems and outperformed human-designed and tuned optimizers. Our work is set to supply a brand-new perspective to L2O research. Codes are available at: https: //github.com/VITA-Group/Symbolic-Learning-To-Optimize .
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper11
- A Closer Look at Learned Optimization: Stability, Robustness, and Inductive BiasesJames Harrison, Luke Metz, Jascha Sohl-DicksteinNeurIPS 2022 · 被引用 41 次
- Outline, Then Details: Syntactically Guided Coarse-To-Fine Code GenerationWenqing Zheng, S. P. Sharan, Ajay Kumar Jaiswal, Kevin Wang 等ICML 2023 · 被引用 35 次
- SYMBOL: Generating Flexible Black-Box Optimizers through Symbolic Equation LearningJiacheng Chen, Zeyuan Ma, Hongshu Guo, Yining Ma 等ICLR 2024 · 被引用 27 次
- Symbolic Distillation for Learned TCP Congestion ControlS. P. Sharan, Wenqing Zheng, Kuo-Feng Hsu, Jiarong Xing 等NeurIPS 2022 · 被引用 9 次
- Efficient Non-Parametric Optimizer Search for Diverse TasksRuochen Wang, Yuanhao Xiong, Minhao Cheng, Cho-Jui HsiehNeurIPS 2022 · 被引用 7 次
它引用的顶会 Paper13
- Discovering Symbolic Models from Deep Learning with Inductive BiasesMiles D. Cranmer, Alvaro Sanchez-Gonzalez, Peter W. Battaglia, Rui Xu 等NeurIPS 2020 · 被引用 736 次
- AutoML-Zero: Evolving Machine Learning Algorithms From ScratchEsteban Real, Chen Liang, David R. So, Quoc V. LeICML 2020 · 被引用 265 次
- Anti-Oversmoothing in Deep Vision Transformers via the Fourier Domain Analysis: From Theory to PracticePeihao Wang, Wenqing Zheng, Tianlong Chen, Zhangyang WangICLR 2022 · 被引用 212 次
- RNA Secondary Structure Prediction By Learning Unrolled AlgorithmsXinshi Chen, Yu Li, Ramzan Umarov, Xin Gao 等ICLR 2020 · 被引用 134 次
- Self-PU: Self Boosted and Calibrated Positive-Unlabeled TrainingXuxi Chen, Wuyang Chen, Tianlong Chen, Ye Yuan 等ICML 2020 · 被引用 100 次
相关 Paper
- Training Stronger Baselines for Learning to OptimizeTianlong Chen, Weiyi Zhang, Jingyang Zhou, Shiyu Chang 等NeurIPS 2020 · 被引用 61 次
- M-L2O: Towards Generalizable Learning-to-Optimize by Test-Time Fast Self-AdaptationJunjie Yang, Xuxi Chen, Tianlong Chen, Zhangyang Wang 等ICLR 2023
- Towards Constituting Mathematical Structures for Learning to OptimizeJialin Liu, Xiaohan Chen, Zhangyang Wang, Wotao Yin 等ICML 2023 · 被引用 18 次
- μLO: Compute-Efficient Meta-Generalization of Learned OptimizersBenjamin Thérien, Charles-Étienne Joseph, Boris Knyazev, Edouard Oyallon 等ICLR 2026 · 被引用 10 次
- Celo2: Towards Learned Optimization Free LunchAbhinav Moudgil, Boris Knyazev, Eugene BelilovskyICLR 2026 · 被引用 1 次
