NOVAS: Non-convex Optimization via Adaptive Stochastic Search for End-to-end Learning and Control
Ioannis Exarchos, Marcus Aloysius Pereira, Ziyi Wang, Evangelos A. Theodorou
摘要
In this work we propose the use of adaptive stochastic search as a building block for general, non-convex optimization operations within deep neural network architectures. Specifically, for an objective function located at some layer in the network and parameterized by some network parameters, we employ adaptive stochastic search to perform optimization over its output. This operation is differentiable and does not obstruct the passing of gradients during backpropagation, thus enabling us to incorporate it as a component in end-to-end learning. We study the proposed optimization module's properties and benchmark it against two existing alternatives on a synthetic energy-based structured prediction task, and further showcase its use in stochastic optimal control applications. * Equal contribution. 1 To distinguish between the optimization of the entire network as opposed to that of the optimization module, we frequently refer to the former as global or outer-loop optimization and to the latter as local or inner-loop optimization.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper2
相关 Paper
- Learning with Differentiable Pertubed OptimizersQuentin Berthet, Mathieu Blondel, Olivier Teboul, Marco Cuturi 等NeurIPS 2020 · 被引用 181 次
- NAS-OoD: Neural Architecture Search for Out-of-Distribution GeneralizationHaoyue Bai, Fengwei Zhou, Lanqing Hong, Nanyang Ye 等ICCV 2021 · 被引用 46 次
- Adversarial Localized Energy Network for Structured PredictionPingbo Pan, Ping Liu, Yan Yan, Tianbao Yang 等AAAI 2020 · 被引用 9 次
- iDARTS: Differentiable Architecture Search with Stochastic Implicit GradientsMiao Zhang, Steven W. Su, Shirui Pan, Xiaojun Chang 等ICML 2021 · 被引用 81 次
- Unchain the Search Space with Hierarchical Differentiable Architecture SearchGuanting Liu, Yujie Zhong, Sheng Guo, Matthew R. Scott 等AAAI 2021 · 被引用 3 次
