Provably Efficient Online Hyperparameter Optimization with Population-Based Bandits
Jack Parker-Holder, Vu Nguyen, Stephen J. Roberts
摘要
Many of the recent triumphs in machine learning are dependent on well-tuned hyperparameters. This is particularly prominent in reinforcement learning (RL) where a small change in the configuration can lead to failure. Despite the importance of tuning hyperparameters, it remains expensive and is often done in a naive and laborious way. A recent solution to this problem is Population Based Training (PBT) which updates both weights and hyperparameters in a single training run of a population of agents. PBT has been shown to be particularly effective in RL, leading to widespread use in the field. However, PBT lacks theoretical guarantees since it relies on random heuristics to explore the hyperparameter space. This inefficiency means it typically requires vast computational resources, which is prohibitive for many small and medium sized labs. In this work, we introduce the first provably efficient PBT-style algorithm, Population-Based Bandits (PB2). PB2 uses a probabilistic model to guide the search in an efficient way, making it possible to discover high performing hyperparameter configurations with far fewer agents than typically required by PBT. We show in a series of RL experiments that PB2 is able to achieve high performance with a modest computational budget.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper20
- Hyperparameters in Reinforcement Learning and How To Tune ThemTheresa Eimer, Marius Lindauer, Roberta RaileanuICML 2023 · 被引用 96 次
- Revisiting Design Choices in Offline Model Based Reinforcement LearningCong Lu, Philip J. Ball, Jack Parker-Holder, Michael A. Osborne 等ICLR 2022 · 被引用 65 次
- Sample-Efficient Automated Deep Reinforcement LearningJörg K. H. Franke, Gregor Köhler, André Biedenkapp, Frank HutterICLR 2021 · 被引用 49 次
- Optimal Transport Kernels for Sequential and Parallel Neural Architecture SearchVu Nguyen, Tam Le, Makoto Yamada, Michael A. OsborneICML 2021 · 被引用 42 次
- Bayesian Optimization for Iterative LearningVu Nguyen, Sebastian Schulze, Michael A. OsborneNeurIPS 2020 · 被引用 38 次
它引用的顶会 Paper3
- Bayesian Optimisation over Multiple Continuous and Categorical InputsBin Xin Ru, Ahsan S. Alvi, Vu Nguyen, Michael A. Osborne 等ICML 2020 · 被引用 119 次
- Knowing The What But Not The Where in Bayesian OptimizationVu Nguyen, Michael A. OsborneICML 2020 · 被引用 42 次
- Bayesian Optimization for Iterative LearningVu Nguyen, Sebastian Schulze, Michael A. OsborneNeurIPS 2020 · 被引用 38 次
相关 Paper
- Tuning Mixed Input Hyperparameters on the Fly for Efficient Population Based AutoRLJack Parker-Holder, Vu Nguyen, Shaan Desai, Stephen J. RobertsNeurIPS 2021 · 被引用 22 次
- Iterated Population Based Training with Task-Agnostic RestartsAlexander Chebykin, Tanja Alderliesten, Peter A.N BosmanICML 2026
- Multi-Objective Population Based TrainingArkadiy Dushatskiy, Alexander Chebykin, Tanja Alderliesten, Peter A. N. BosmanICML 2023 · 被引用 4 次
- Accelerating and Improving AlphaZero Using Population Based TrainingTi-Rong Wu, Ting-Han Wei, I-Chen WuAAAI 2020 · 被引用 19 次
- Fast Population-Based Reinforcement Learning on a Single MachineArthur Flajolet, Claire Bizon Monroc, Karim Beguir, Thomas PierrotICML 2022 · 被引用 11 次
