Structure Detection for Contextual Reinforcement Learning
Tianyue Zhou, Jung-Hoon Cho, Cathy Wu
摘要
Contextual Reinforcement Learning (CRL) tackles the problem of solving a set of related Contextual Markov Decision Processes (CMDPs) that vary across different context variables. Traditional approaches---independent training and multi-task learning---struggle with either excessive computational costs or negative transfer. A recently proposed multi-policy approach, Model-Based Transfer Learning (MBTL), has demonstrated effectiveness by strategically selecting a few tasks to train and zero-shot transfer. However, CMDPs encompass a wide range of problems, exhibiting structural properties that vary from problem to problem. As such, different task selection strategies are suitable for different CMDPs. In this work, we introduce Structure Detection MBTL (SD-MBTL), a generic framework that dynamically identifies the underlying generalization structure of CMDP and selects an appropriate MBTL algorithm. For instance, we observe Mountain structure in which generalization performance degrades from the training performance of the target task as the context difference increases. We thus propose M/GP-MBTL, which detects the structure and adaptively switches between a Gaussian Process-based approach and a clustering-based approach. Extensive experiments on synthetic data and CRL benchmarks—covering continuous control, traffic control, and agricultural management—show that M/GP-MBTL surpasses the strongest prior method by 12.49% on the aggregated metric. These results highlight the promise of online structure detection for guiding source task selection in complex CRL environments.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper10
- Which Tasks Should Be Learned Together in Multi-task Learning?Trevor Standley, Amir Zamir, Dawn Chen, Leonidas J. Guibas 等ICML 2020 · 被引用 651 次
- Agent57: Outperforming the Atari Human BenchmarkAdrià Puigdomènech Badia, Bilal Piot, Steven Kapturowski, Pablo Sprechmann 等ICML 2020 · 被引用 584 次
- Multi-Task Reinforcement Learning with Context-based RepresentationsShagun Sodhani, Amy Zhang, Joelle PineauICML 2021 · 被引用 241 次
- PaCo: Parameter-Compositional Multi-task Reinforcement LearningLingfeng Sun, Haichao Zhang, Wei Xu, Masayoshi TomizukaNeurIPS 2022 · 被引用 72 次
- Multi-Task Reinforcement Learning with Mixture of Orthogonal ExpertsAhmed Hendawy, Jan Peters, Carlo D'EramoICLR 2024 · 被引用 45 次
相关 Paper
- Model-Based Transfer Learning for Contextual Reinforcement LearningJung-Hoon Cho, Vindula Jayawardana, Sirui Li, Cathy WuNeurIPS 2024 · 被引用 14 次
- Learning Robust State Abstractions for Hidden-Parameter Block MDPsAmy Zhang, Shagun Sodhani, Khimya Khetarpal, Joelle PineauICLR 2021 · 被引用 5 次
- Trajectory-wise Multiple Choice Learning for Dynamics Generalization in Reinforcement LearningYounggyo Seo, Kimin Lee, Ignasi Clavera Gilaberte, Thanard Kurutach 等NeurIPS 2020 · 被引用 51 次
- CrossLight: Offline-to-Online Reinforcement Learning for Cross-City Traffic Signal ControlQian Sun, Rui Zha, Le Zhang, Jingbo Zhou 等KDD 2024 · 被引用 9 次
- Towards an Information Theoretic Framework of Context-Based Offline Meta-Reinforcement LearningLanqing Li, Hai Zhang, Xinyu Zhang, Shatong Zhu 等NeurIPS 2024 · 被引用 24 次
