Creating Multi-Level Skill Hierarchies in Reinforcement Learning
Joshua B. Evans, Özgür Simsek
摘要
What is a useful skill hierarchy for an autonomous agent? We propose an answer based on a graphical representation of how the interaction between an agent and its environment may unfold. Our approach uses modularity maximisation as a central organising principle to expose the structure of the interaction graph at multiple levels of abstraction. The result is a collection of skills that operate at varying time scales, organised into a hierarchy, where skills that operate over longer time scales are composed of skills that operate over shorter time scales. The entire skill hierarchy is generated automatically, with no human intervention, including the skills themselves (their behaviour, when they can be called, and when they terminate) as well as the hierarchical dependency structure between them. In a wide range of environments, this approach generates skill hierarchies that are intuitively appealing and that considerably improve the learning performance of the agent.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper7
- Structural Information-based Hierarchical Diffusion for Offline Reinforcement LearningXianghua Zeng, Hao Peng, Yicheng Pan, Angsheng Li 等NeurIPS 2025 · 被引用 4 次
- Skill-Driven Neurosymbolic State AbstractionsAlper Ahmetoglu, Steven James, Cameron Allen, Sam Lobel 等NeurIPS 2025 · 被引用 3 次
- Novel Exploration via OrthogonalityAndreas Theophilou, Özgür SimsekNeurIPS 2025 · 被引用 1 次
- The Cost of Commitment in Option-Based Hierarchical RLRandy Lefebvre, Audrey DurandICML 2026
- Accelerating Task Generalisation with Multi-Level Skill HierarchiesThomas P. Cannon, Özgür SimsekICLR 2025
相关 Paper
- Skill Discovery for Exploration and Planning using Deep Skill GraphsAkhil Bagaria, Jason K. Senthil, George KonidarisICML 2021 · 被引用 73 次
- When Do Skills Help Reinforcement Learning? A Theoretical Analysis of Temporal AbstractionsZhening Li, Gabriel Poesia, Armando Solar-LezamaICML 2024 · 被引用 1 次
- Unsupervised Hierarchical Skill DiscoveryDamion Harvey, Geraud Nangue Tasse, Benjamin Rosman, Branden Ingram 等ICML 2026 · 被引用 1 次
- Possibility Before Utility: Learning And Using Hierarchical AffordancesRobby Costales, Shariq Iqbal, Fei ShaICLR 2022 · 被引用 5 次
- Option Discovery using Deep Skill ChainingAkhil Bagaria, George KonidarisICLR 2020 · 被引用 126 次
