Lifelong Learning with a Changing Action Set
Yash Chandak, Georgios Theocharous, Chris Nota, Philip S. Thomas
Abstract
In many real-world sequential decision making problems, the number of available actions (decisions) can vary over time. While problems like catastrophic forgetting, changing transition dynamics, changing rewards functions, etc. have been well-studied in the lifelong learning literature, the setting where the size of the action set changes remains unaddressed. In this paper, we present first steps towards developing an algorithm that autonomously adapts to an action set whose size changes over time. To tackle this open problem, we break it into two problems that can be solved iteratively: inferring the underlying, unknown, structure in the space of actions and optimizing a policy that leverages this structure. We demonstrate the efficiency of this approach on large-scale real-world lifelong learning problems.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 30cfd13f-c08e-4e96-9ff0-543d984c4b25Cited by top-tier papers10
- Parameterizing Branch-and-Bound Search Trees to Learn Branching PoliciesGiulia Zarpellon, Jason Jo, Andrea Lodi, Yoshua BengioAAAI 2021 · 123 citations
- Towards Safe Policy Improvement for Non-Stationary MDPsYash Chandak, Scott M. Jordan, Georgios Theocharous, Martha White et al.NeurIPS 2020 · 47 citations
- Generalization to New Actions in Reinforcement LearningAyush Jain, Andrew Szot, Joseph J. LimICML 2020 · 39 citations
- In-Context Reinforcement Learning for Variable Action SpacesViacheslav Sinii, Alexander Nikulin, Vladislav Kurenkov, Ilya Zisman et al.ICML 2024 · 26 citations
- Perceiving the World: Question-guided Reinforcement Learning for Text-based GamesYunqiu Xu, Meng Fang, Ling Chen, Yali Du et al.ACL 2022 · 22 citations
Related papers
- Deep Reinforcement Learning amidst Continual Structured Non-StationarityAnnie Xie, James Harrison, Chelsea FinnICML 2021 · 43 citations
- Sleeping Reinforcement LearningSimone Drago, Marco Mussi, Alberto Maria MetelliICML 2025
- CORDS - Continuous Representations of Discrete StructuresTin Hadži Veljković, Erik J Bekkers, Michael Tiemann, Jan-Willem van de MeentICLR 2026 · 1 citation
- Provably Efficient Causal Model-Based Reinforcement Learning for Systematic GeneralizationMirco Mutti, Riccardo De Santi, Emanuele Rossi, Juan Felipe Calderón et al.AAAI 2023 · 17 citations
- Lifelong Policy Gradient Learning of Factored Policies for Faster Training Without ForgettingJorge A. Mendez, Boyu Wang, Eric EatonNeurIPS 2020 · 42 citations
