DHRL: A Graph-Based Approach for Long-Horizon and Sparse Hierarchical Reinforcement Learning
Seungjae Lee, Jigang Kim, Inkyu Jang, H. Jin Kim
摘要
Hierarchical Reinforcement Learning (HRL) has made notable progress in complex control tasks by leveraging temporal abstraction. However, previous HRL algorithms often suffer from serious data inefficiency as environments get large. The extended components, , goal space and length of episodes, impose a burden on either one or both high-level and low-level policies since both levels share the total horizon of the episode. In this paper, we present a method of Decoupling Horizons Using a Graph in Hierarchical Reinforcement Learning (DHRL) which can alleviate this problem by decoupling the horizons of high-level and low-level policies and bridging the gap between the length of both horizons using a graph. DHRL provides a freely stretchable high-level action interval, which facilitates longer temporal abstraction and faster training in complex tasks. Our method outperforms state-of-the-art HRL algorithms in typical HRL environments. Moreover, DHRL achieves long and complex locomotion and manipulation tasks.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper9
- CQM: Curriculum Reinforcement Learning with a Quantized World ModelSeungjae Lee, Daesol Cho, Jonghae Park, H. Jin KimNeurIPS 2023 · 被引用 18 次
- Breadth-First Exploration on Adaptive Grid for Reinforcement LearningYoungsik Yoon, Gangbok Lee, Sungsoo Ahn, Jungseul OkICML 2024 · 被引用 5 次
- Strict Subgoal Execution: Reliable Long-Horizon Planning in Hierarchical Reinforcement LearningSeungyul Han, Jaebak Hwang, Sanghyeon Lee, Jeongmo KimICLR 2026 · 被引用 3 次
- Latent State-Predictive Exploration for Deep Reinforcement LearningYiming Wang, Kaiyan Zhao, Borong Zhang, Yan Li 等AAAI 2026 · 被引用 1 次
- Wavelet Policy: Lifting Scheme for Policy Learning in Long-Horizon TasksHao Huang, Shuaihang Yuan, Geeta Chandra Raju Bethala, Congcong Wen 等ICCV 2025
它引用的顶会 Paper5
- Skew-Fit: State-Covering Self-Supervised Reinforcement LearningVitchyr Pong, Murtaza Dalal, Steven Lin, Ashvin Nair 等ICML 2020 · 被引用 303 次
- Generating Adjacency-Constrained Subgoals in Hierarchical Reinforcement LearningTianren Zhang, Shangqi Guo, Tian Tan, Xiaolin Hu 等NeurIPS 2020 · 被引用 112 次
- World Model as a Graph: Learning Latent Landmarks for PlanningLunjun Zhang, Ge Yang, Bradly C. StadieICML 2021 · 被引用 90 次
- Landmark-Guided Subgoal Generation in Hierarchical Reinforcement LearningJunsu Kim, Younggyo Seo, Jinwoo ShinNeurIPS 2021 · 被引用 90 次
- Skill Discovery for Exploration and Planning using Deep Skill GraphsAkhil Bagaria, Jason K. Senthil, George KonidarisICML 2021 · 被引用 73 次
相关 Paper
- Reconciling Spatial and Temporal Abstractions for Goal RepresentationMehdi Zadem, Sergio Mover, Sao Mai NguyenICLR 2024 · 被引用 8 次
- Hierarchical Reinforcement Learning with Targeted Causal InterventionsMohammadsadegh Khorasani, Saber Salehkaleybar, Negar Kiyavash, Matthias GrossglauserICML 2025
- Value Function Spaces: Skill-Centric State Abstractions for Long-Horizon ReasoningDhruv Shah, Peng Xu, Yao Lu, Ted Xiao 等ICLR 2022 · 被引用 50 次
- Enhancing Exploration and Exploitation in Hierarchical Reinforcement Learning with Subgoal Graph LearningYibo Zhang, Dengpeng XingAAAI 2026
- State-Conditioned Adversarial Subgoal GenerationVivienne Huiling Wang, Joni Pajarinen, Tinghuai Wang, Joni-Kristian KämäräinenAAAI 2023 · 被引用 16 次
