Learning Achievement Structure for Structured Exploration in Domains with Sparse Reward
Zihan Zhou, Animesh Garg
Abstract
We propose Structured Exploration with Achievements (SEA), a multi-stage reinforcement learning algorithm designed for achievement-based environments, a particular type of environment with an internal achievement set. SEA first uses offline data to learn a representation of the known achievements with a determinant loss function, then recovers the dependency graph of the learned achievements with a heuristic algorithm, and finally interacts with the environment online to learn policies that master known achievements and explore new ones with a controller built with the recovered dependency graph. We empirically demonstrate that SEA can recover the achievement structure accurately and improve exploration in hard domains such as Crafter that are procedurally generated with high-dimensional observations like images.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 33acab37-2243-4ab7-a4b0-3db9d98cbb3dCited by top-tier papers1
Ask how each one uses itBuilds on8
- Mastering Atari with Discrete World ModelsDanijar Hafner, Timothy P. Lillicrap, Mohammad Norouzi, Jimmy BaICLR 2021 · 1,170 citations
- Never Give Up: Learning Directed Exploration StrategiesAdrià Puigdomènech Badia, Pablo Sprechmann, Alex Vitvitskyi, Zhaohan Daniel Guo et al.ICLR 2020 · 349 citations
- Effective Diversity in Population Based Reinforcement LearningJack Parker-Holder, Aldo Pacchiano, Krzysztof Marcin Choromanski, Stephen J. RobertsNeurIPS 2020 · 195 citations
- Benchmarking the Spectrum of Agent CapabilitiesDanijar HafnerICLR 2022 · 193 citations
- Dynamical Distance Learning for Semi-Supervised and Unsupervised Skill DiscoveryKristian Hartikainen, Xinyang Geng, Tuomas Haarnoja, Sergey LevineICLR 2020 · 94 citations
Related papers
- Investigating the Role of Model-Based Learning in Exploration and TransferJacob C. Walker, Eszter Vértes, Yazhe Li, Gabriel Dulac-Arnold et al.ICML 2023 · 8 citations
- Unsupervised Hierarchical Skill DiscoveryDamion Harvey, Geraud Nangue Tasse, Benjamin Rosman, Branden Ingram et al.ICML 2026 · 1 citation
- SPRING: Studying Papers and Reasoning to play GamesYue Wu, So Yeon Min, Shrimai Prabhumoye, Yonatan Bisk et al.NeurIPS 2023 · 32 citations
- Craftax: A Lightning-Fast Benchmark for Open-Ended Reinforcement LearningMichael T. Matthews, Michael Beukman, Benjamin Ellis, Mikayel Samvelyan et al.ICML 2024 · 71 citations
- SEER: Facilitating Structured Reasoning and Explanation via Reinforcement LearningGuoxin Chen, Kexin Tang, Chao Yang, Fuying Ye et al.ACL 2024 · 5 citations
