Generalizing Gaussian Smoothing for Random Search
Katelyn Gao, Ozan Sener
摘要
Gaussian smoothing (GS) is a derivative-free optimization (DFO) algorithm that estimates the gradient of an objective using perturbations of the current parameters sampled from a standard normal distribution. We generalize it to sampling perturbations from a larger family of distributions. Based on an analysis of DFO for non-convex functions, we propose to choose a distribution for perturbations that minimizes the mean squared error (MSE) of the gradient estimate. We derive three such distributions with provably smaller MSE than Gaussian smoothing. We conduct evaluations of the three sampling distributions on linear regression, reinforcement learning, and DFO benchmarks in order to validate our claims. Our proposal improves on GS with the same computational complexity, and are usually competitive with and often outperform Guided ES (Maheswaranathan et al., 2019) and Orthogonal ES (Choromanski et al., 2018) , two computationally more expensive algorithms that adapt the covariance matrix of normally distributed perturbations.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper9
- Continuation Path Learning for Homotopy OptimizationXi Lin, Zhiyuan Yang, Xiaoyuan Zhang, Qingfu ZhangICML 2023 · 被引用 18 次
- Variance-Reduced Gradient Estimation via Noise-Reuse in Online Evolution StrategiesOscar Li, James Harrison, Jascha Sohl-Dickstein, Virginia Smith 等NeurIPS 2023 · 被引用 11 次
- Adaptive Stochastic Gradient Algorithm for Black-box Multi-Objective LearningFeiyang Ye, Yueming Lyu, Xuehao Wang, Yu Zhang 等ICLR 2024 · 被引用 5 次
- Learning a Zeroth-Order Optimizer for Fine-Tuning LLMsKairun Zhang, Haoyu Li, Yanjun Zhao, Yifan Sun 等ICML 2026 · 被引用 1 次
- Sharpness-Aware Black-Box OptimizationFeiyang Ye, Yueming Lyu, Xuehao Wang, Masashi Sugiyama 等ICLR 2025
它引用的顶会 Paper2
相关 Paper
- Dynamic Anisotropic Smoothing for Noisy Derivative-Free OptimizationSam Reifenstein, Timothée G. Leleu, Yoshihisa YamamotoICML 2024 · 被引用 3 次
- On the Optimal Construction of Unbiased Gradient Estimators for Zeroth-Order OptimizationShaocong Ma, Heng HuangNeurIPS 2025 · 被引用 4 次
- DoWG Unleashed: An Efficient Universal Parameter-Free Gradient Descent MethodAhmed Khaled, Konstantin Mishchenko, Chi JinNeurIPS 2023 · 被引用 49 次
- Black-Box Generalization: Stability of Zeroth-Order LearningKonstantinos E. Nikolakakis, Farzin Haddadpour, Dionysios S. Kalogerias, Amin KarbasiNeurIPS 2022
- Guided Zeroth-Order Methods for Stochastic Non-convex Problems with Decision-Dependent DistributionsYuya Hikima, Hiroshi Sawada, Akinori FujinoICML 2025
