Sample Complexity of Goal-Conditioned Hierarchical Reinforcement Learning
Arnaud Robert, Ciara Pike-Burke, Aldo A. Faisal
Abstract
Hierarchical Reinforcement Learning (HRL) algorithms can perform planning at multiple levels of abstraction. Empirical results have shown that state or temporal abstractions might significantly improve the sample efficiency of algorithms. Yet, we still do not have a complete understanding of the basis of those efficiency gains, nor any theoretically-grounded design rules. In this paper, we derive a lower bound on the sample complexity for the considered class of goal-conditioned HRL algorithms. The proposed lower bound empowers us to quantify the benefits of hierarchical decomposition and leads to the design of a simple Q-learning-type algorithm that leverages hierarchical decompositions. We empirically validate our theoretical findings by investigating the sample complexity of the proposed hierarchical algorithm on a spectrum of tasks (hierarchical n -rooms, Gymnasium’s Taxi). The hierarchical n -rooms tasks were designed to allow us to dial their complexity over multiple orders of magnitude. Our theory and algorithmic findings provide a step towards answering the foundational question of quantifying the improvement hierarchical decomposition offers over monolithic solutions in reinforcement learning.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext eb842c5e-9aa2-466b-b776-3473394b40f2Cited by top-tier papers2
- Self-Improving Skill Learning for Robust Skill-based Meta-Reinforcement LearningSeungyul Han, Sanghyeon Lee, Sangjun Bae, Yisak ParkICLR 2026 · 5 citations
- A theoretical case-study of Scalable Oversight in Hierarchical Reinforcement LearningTom Yan, Zachary C. LiptonNeurIPS 2024 · 3 citations
Builds on1
Related papers
- DHRL: A Graph-Based Approach for Long-Horizon and Sparse Hierarchical Reinforcement LearningSeungjae Lee, Jigang Kim, Inkyu Jang, H. Jin KimNeurIPS 2022 · 33 citations
- Hierarchical Reinforcement Learning with Timed SubgoalsNico Gürtler, Dieter Büchler, Georg MartiusNeurIPS 2021 · 43 citations
- Globally Optimal Hierarchical Reinforcement Learning for Linearly-Solvable Markov Decision ProcessesGuillermo Infante, Anders Jonsson, Vicenç GómezAAAI 2022 · 8 citations
- HIQL: Offline Goal-Conditioned RL with Latent States as ActionsSeohong Park, Dibya Ghosh, Benjamin Eysenbach, Sergey LevineNeurIPS 2023 · 173 citations
- Reconciling Spatial and Temporal Abstractions for Goal RepresentationMehdi Zadem, Sergio Mover, Sao Mai NguyenICLR 2024 · 8 citations
