Variance-Reduced Gradient Estimation via Noise-Reuse in Online Evolution Strategies
Oscar Li, James Harrison, Jascha Sohl-Dickstein, Virginia Smith, Luke Metz
摘要
Unrolled computation graphs are prevalent throughout machine learning but present challenges to automatic differentiation (AD) gradient estimation methods when their loss functions exhibit extreme local sensitivtiy, discontinuity, or blackbox characteristics. In such scenarios, online evolution strategies methods are a more capable alternative, while being more parallelizable than vanilla evolution strategies (ES) by interleaving partial unrolls and gradient updates. In this work, we propose a general class of unbiased online evolution strategies methods. We analytically and empirically characterize the variance of this class of gradient estimators and identify the one with the least variance, which we term Noise-Reuse Evolution Strategies (NRES). Experimentally 3 , we show NRES results in faster convergence than existing AD and ES methods in terms of wall-clock time and number of unroll steps across a variety of applications, including learning dynamical systems, meta-training learned optimizers, and reinforcement learning.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper3
- Evolution Strategies at the HyperscaleBidipta Sarkar, Mattie Fellows, Juan Duque, Alistair Letcher 等ICML 2026 · 被引用 16 次
- μLO: Compute-Efficient Meta-Generalization of Learned OptimizersBenjamin Thérien, Charles-Étienne Joseph, Boris Knyazev, Edouard Oyallon 等ICLR 2026 · 被引用 10 次
- Neural Evolution Strategy for Black-box Pareto Set LearningChengyu Lu, Zhenhua Li, Xi Lin, Ji Cheng 等NeurIPS 2025
它引用的顶会 Paper6
- Dataset Distillation by Matching Training TrajectoriesGeorge Cazenavette, Tongzhou Wang, Antonio Torralba, Alexei A. Efros 等CVPR 2022 · 被引用 198 次
- Unbiased Gradient Estimation in Unrolled Computation Graphs with Persistent Evolution StrategiesPaul Vicol, Luke Metz, Jascha Sohl-DicksteinICML 2021 · 被引用 77 次
- Learning by Directional Gradient DescentDavid Silver, Anirudh Goyal, Ivo Danihelka, Matteo Hessel 等ICLR 2022 · 被引用 44 次
- A Closer Look at Learned Optimization: Stability, Robustness, and Inductive BiasesJames Harrison, Luke Metz, Jascha Sohl-DicksteinNeurIPS 2022 · 被引用 41 次
- Generalizing Gaussian Smoothing for Random SearchKatelyn Gao, Ozan SenerICML 2022 · 被引用 22 次
相关 Paper
- Low-Variance Gradient Estimation in Unrolled Computation Graphs with ES-SinglePaul VicolICML 2023 · 被引用 8 次
- Learning Discrete Structured Variational Auto-Encoder using Natural Evolution StrategiesAlon Berliner, Guy Rotman, Yossi Adi, Roi Reichart 等ICLR 2022 · 被引用 5 次
- EvoGrad: Evolutionary-Weighted Gradient and Hessian Learning for Black-Box OptimizationYedidya Kfir, Elad Sarafian, Yoram Louzoun, Sarit KrausAAAI 2026
- Explicit Gradient Learning for Black-Box OptimizationElad Sarafian, Mor Sinay, Yoram Louzoun, Noa Agmon 等ICML 2020
- Discovering Evolution Strategies via Meta-Black-Box OptimizationRobert Tjarko Lange, Tom Schaul, Yutian Chen, Tom Zahavy 等ICLR 2023 · 被引用 21 次
