Information-theoretic Task Selection for Meta-Reinforcement Learning
Ricardo Luna Gutiérrez, Matteo Leonetti
Abstract
In Meta-Reinforcement Learning (meta-RL) an agent is trained on a set of tasks to prepare for and learn faster in new, unseen, but related tasks. The training tasks are usually hand-crafted to be representative of the expected distribution of test tasks and hence all used in training. We show that given a set of training tasks, learning can be both faster and more effective (leading to better performance in the test tasks), if the training tasks are appropriately selected. We propose a task selection algorithm, Information-Theoretic Task Selection (ITTS), based on information theory, which optimizes the set of tasks used for training in meta-RL, irrespectively of how they are generated. The algorithm establishes which training tasks are both sufficiently relevant for the test tasks, and different enough from one another. We reproduce different meta-RL experiments from the literature and show that ITTS improves the final performance in all of them.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext ddd93051-bf1b-4431-8bde-76c78cf6f67cCited by top-tier papers7
- Offline Meta-Reinforcement Learning with Advantage WeightingEric Mitchell, Rafael Rafailov, Xue Bin Peng, Sergey Levine et al.ICML 2021 · 122 citations
- Meta-learning with an Adaptive Task SchedulerHuaxiu Yao, Yu Wang, Ying Wei, Peilin Zhao et al.NeurIPS 2021 · 61 citations
- Learning to Adapt via Latent Domains for Adaptive Semantic SegmentationYunan Liu, Shanshan Zhang, Yang Li, Jian YangNeurIPS 2021 · 20 citations
- Meta-Learning with Neural Bandit SchedulerYunzhe Qi, Yikun Ban, Tianxin Wei, Jiaru Zou et al.NeurIPS 2023 · 14 citations
- Multidimensional Belief Quantification for Label-Efficient Meta-LearningDeep Shankar Pandey, Qi YuCVPR 2022 · 8 citations
Related papers
- Meta-Q-LearningRasool Fakoor, Pratik Chaudhari, Stefano Soatto, Alexander J. SmolaICLR 2020 · 162 citations
- Offline Meta-Reinforcement Learning with Online Self-SupervisionVitchyr H. Pong, Ashvin Nair, Laura Smith, Catherine Huang et al.ICML 2022 · 78 citations
- Distributionally Adaptive Meta Reinforcement LearningAnurag Ajay, Abhishek Gupta, Dibya Ghosh, Sergey Levine et al.NeurIPS 2022 · 21 citations
- Improving Generalization in Meta Reinforcement Learning using Learned ObjectivesLouis Kirsch, Sjoerd van Steenkiste, Jürgen SchmidhuberICLR 2020 · 132 citations
- Improving Generalization in Meta-RL with Imaginary Tasks from Latent Dynamics MixtureSuyoung Lee, Sae-Young ChungNeurIPS 2021 · 23 citations
